- Photo
- —
- Name
- —
- Authors
- Jitendra Malik, Serge Belongie, Thomas Leung, Jianbo Shi
- Year
- 1999
- Field
- Computer vision
- One-line gist
- Pushes beyond raw pixels to represent images in terms of objects and segments, laying groundwork for modern vision models.
- Why I care
- It’s a reminder that all the fancy deep nets are still chasing the core idea: see scenes as objects, not noise.
- Best stat or figure
- The segmentation examples where messy real-world scenes suddenly break into clean, meaningful regions.
- Difficulty
- Medium-hard; you’ll need some comfort with both stats and vision basics.
- Where to read
- Papers from the IEEE or CVPR proceedings; PDFs are widely mirrored.
- My take
- Reading it now feels like seeing the early sketches for the stuff your phone casually does every time you open the camera.