1. When you say you're built on top of ROS, do you mean the autonomy stack you'd deploy in an actual robot is built on ROS? Are you using ROS or ROS2?
2. What hardware does your current autonomy stack use? For parts of your stack that'd depend on using deep learning based methods (e.g. any image or lidar data), the models that you'd train would have to use a lot of collected and annotated data specific to a particular problem/industry and this would take a non-trivial amount of time, especially since you said the camera/sensor configuration is not fixed can can potentially be decided I the simulator by the user. How do you plan to tackle this?
PS: Any potential internship opportunities at Polymath in the coming few months that I can apply to?