Thanks for sharing @jonbaer! I’m one of the co-founders of Roboflow. Some additi...

riedel · on Dec 31, 2022

Might be good not only to show that YOLO and other do not fully generalize to the presented image domains but also that some specialized SOTA models e.g. for areal images perform well in the subcategories to actually calibrate the benchmark. I otherwise would wonder about the quality of the labels (e.g being contradictive).

yeldarb · on Dec 31, 2022

Good idea. I haven’t looked too closely yet at the “hard” datasets.

We originally considered “fixing” the labels on these datasets by hand, but ultimately decided that label error is one of the challenges “real world” datasets have that models should work to become more robust against. There is some selection bias in that we did make sure that the datasets we chose passed the eye test (in other words, it looked like the user spent a considerable amount of time annotating & a sample of the images looked like they labeled some object of interest).

For aerial images in particular my guess would be that these models suffer from the “small object problem”[1] where the subjects are tiny compared to the size of the image. Trying a sliding window based approach like SAHI[2] on them would probably produce much better results (at the expense of much lower inference speed).

[1] https://blog.roboflow.com/detect-small-objects/

[2] https://github.com/obss/sahi

6gvONxR4sf7o · on Dec 31, 2022

It’s at least worth fixing the test sets, so that the test metrics measure what they ought to.

throwaway20222 · on Dec 31, 2022

I have experimented with the platform and am a huge fan. Thanks for all you are doing - have you thought about or are you integrated with tools like Synthesis AI or Sundial.ai? They don’t cover all of the wide range of data I want to gather, but make it so easy for fixed objects.

yeldarb · on Jan 1, 2023

Haven't heard of those two, but would be really awesome to see an integration. We have an open API[1] for just this reason: we really want to make it easy to use (and source) your data across all the different tools out there. We've recently launched integrations with other labeling[2] and AutoML[3] tools (and have integrations with the big-cloud AutoML tools as well[4]). We're hoping to have a bunch more integrations with other developer tools, labeling services, and MLOps tools & platforms in 2023.

Re synthetic data specifically, we've written a couple of how-to guides for creating data from context augmentation[5], Unity Perception[6], and Stable Diffusion[7] & are talking to some others as well; it seems like a natural integration point (and someplace where we don't need to reinvent the wheel).

[1] https://docs.roboflow.com/rest-api

[2] https://github.com/SkalskiP/make-sense/pull/298

[3] https://github.com/ultralytics/yolov5/discussions/10425

[4] https://docs.roboflow.com/train/pro-third-party-training-int...

[5] https://blog.roboflow.com/how-to-create-a-synthetic-dataset-...

[6] https://blog.roboflow.com/unity-perception-synthetic-dataset...

[7] https://blog.roboflow.com/synthetic-data-with-stable-diffusi...