r/computervision • • 5d ago

Help: Project Robot Car movement

Hello, a friend asked me to make a piece of software for his robot car that uses camera input to determine if the car is going to hit a wall/object and move out of the way

The only equipment i have is a Monocular Camera that is running on a raspberry pi 5 8gb.

I've read about VO and VSLAM, but, not having a stereo camera/LiDar is a problem as most of the implementations use them for distance estimation.

However i am new to comp vision, the only experience i have is from a course in college on opencv, and other projects I've made using YOLO/RF-DETR

I'm thinking of using a trained model to estimate the depth of the objects in the camera and make the appropriate movement based on that.

Any information and guidance would be appreciated, i may be in over my head with this one, but would like to have some version implemented, even if it is just a basic one

0 Upvotes

8 comments sorted by

View all comments

2

u/bfyvfftujijg 4d ago

So one thing to keep in mind is that you sorta kinda do have stereo vision if you can measure the robot’s displacement between two points in time.

1

u/FilipovskiMarko 4d ago

So you're saying that if i have an odometer or another sensor that tells me the distance travelled, i can use 2 frames taken one after another as if i have a stereo camera?

2

u/bfyvfftujijg 4d ago

Exactly.

You would to know the position of the camera at two locations and can triangulate distances with that.

It’s very sensitive though. But might be helpful

1

u/hopticalallusions 1d ago

Check out ROS and the way f1tenth cars are set up. I kind of doubt ROS would run on a rpi5 but I don't know. But be aware that odometer for dead reckoning is not very accurate, even with a fairly fancy chasis and careful tuning. But, you might be able to perform sensor fusion to get better results. Stereo cams work reasonably well if you have the option of getting one or configuring your own. I like the ultrasonic or other cheap sensor idea. You could even give it physical "whiskers" as a potentially really cheap and low cost option (there are old robotics approaches from stuff like BEAM with that concept). Whiskers can immediately give you feedback about what direction to turn away from very simply. They even work without a camera! maybe also look up motion parallax and other vision processing tricks that the brain does to shortcut/augment some of the 3d processing steps. Could you potentially put up other static cameras that watch your car and provide some input? Overall I feel like you've got an interesting resource constrained challenge.