TRUSTED BY +25,000 BUSINESSES

The complete data loop

Everything between raw data and a better model.

One tool for the whole loop — curate, annotate, evaluate, export — and a Python API behind every panel in the UI.

Find the samples worth labeling.

Everything you index is embedded automatically.

Filter by metadata, search by text or image, then run a sampling strategy - diverse coverage, deduplication, outlier detection, similarity, or class balancing - to keep only the subset worth labeling.

Automatic embeddings on index
Text and image similarity search
Six sampling strategies, one query
Start for Free

Label and QA your data in one place.

Create, edit and review annotations right where you work, whether that's a quick check or a full QA pass.

Every labeled object gets its own embedding, so you can search and cluster objects the same way you search whole images. Bring in SAM3 or a LightlyTrain model to auto-label the easy cases first.

Annotate and QA in one place
Object-level search
SAM3 and LightlyTrain auto-labeling plugins
Start for Free

Know exactly where your model breaks.

Run an evaluation against ground truth and get per-sample metrics plus a confusion matrix, no separate script required.

Sort by false positives, false negatives or mIoU to jump straight to the samples your model gets wrong.

Detection, classification and segmentation
Confusion matrix built in
Sort by fp, fn or mIoU
Start for Free

Add or vibe-code your plugins.

Install ready-made plugins for SAM3 segmentation, LightlyTrain auto-labeling and video bounding box propagation, or register your own.

A plugin is a Python function with a form, so anything you can script, your whole team can run from the UI.

SAM3 and LightlyTrain, ready to install
Video bounding box propagation
Build your own with the Python API
Start for Free

Work on the same dataset, together.

Invite your team and give each person a role that matches what they should touch.

Viewers browse and export, Labelers tag and annotate, Editors configure the workspace, Admins manage it all.

Role-based access for your team
Viewer, Labeler, Editor, Admin roles
Centrally managed cloud credentials
Start for Free

Get your data out, in the format you need.

Export annotations or filenames straight from the grid, or automate it in Python. Export the full dataset or just a filtered subset by tag or query.

COCO, YOLO, Pascal VOC and more
Full dataset or filtered subset
GUI or Python, your choice
Start for Free
This is LightlyStudio

One Platform. Two audiences.

LightlyStudio connects everyone working on your ML data pipeline - from ML engineers to labelers. Built for technical and non-technical users.

For ML Engineers

Integrate with your existing ML stack, access SDK and API support, and build on open-source standards designed for flexibility and scale.

Easy automation via Python SDK
Technical support & docs
Import & manage data via code

For Labelers & Project Managers

Use intuitive labeling tools with role-based permissions, dataset versioning, and performance tracking to manage annotation workflows at scale.

User-friendly interface
Labelling & QA tools
Team collaboration
Data versioning

We make it easy to migrate your data from Encord, Voxel51, Ultralytics, V7Labs, Roboflow, or other ML tools. Contact us to learn more.

Why LightlyStudio

Open-source principles.
Enterprise-grade security & scale.

LightlyStudio meets enterprise-grade requirements for compliance, extensibility, and deployment flexibility.

ISO 27001 Certification

International standard for information security management systems

Deploy anywhere

Use your dataset to pretrain a model with just a few lines of code.

Available as Open-Source

A computer vision framework for self-supervised learning developed for research.

Open Source

Recommended for students

  • Recommended for students & individuals
  • pip install
  • No customer support
  • No team collaboration

Start 14 days free trial

  • All features as open-source
  • Hosted by Lightly
  • Quick start via UI
  • Dataset management
  • Team collaboration

On-Premise

Deploy on your infrastructure.

  • All features as open-source
  • Deployed on customer infrastracture
  • Dedicated engineering support
  • Dataset management
  • Team collaboration

LightlyStudio in numbers

2M+

images in one dataset
Indexing, embeddings and search stay responsive as the dataset grows.

16 GB

COCO and ImageNet on a laptop
A Rust core keeps ImageNet-scale work on a 16GB MacBook Pro M1

< 1 min

From pip install to browser
LightlyStudio quickstart loads a sample dataset and opens the app.
Tutorials & Guides

Start with a workflow,
not a blank dataset.

Each tutorial walks through a complete, real workflow — from raw data to a trained or evaluated model.

View LightlyStudio Tutorials
What customers say

World-class ML Teams Choose LightlyStudio

LightlyStudio is used and loved by world’s best ML teams

3x

Faster Model Iteration Cycle
Read more

2M+

Curated Images
Read more

Improved performance YOLO-based models
Read more

DINOv3

Object Detection Pipeline Beating Prior Baselines
Read more

2.3M

Frames Processed In 1 Month
Read more

10%

Model Accuracy Improvement
Read more

3D

SSL Training Pipeline with DINOv2
Read more

82%

Precision on Herd Size Estimation
Read more

2x

Model Deployment Efficiency Gains
Read more

36%

Model Accuracy Improvement
Read more

50%

Reduced Retraining Process Time
Read more

80%

Reduction in Annotation Costs
Read more

Ready to Get Started?

Join 100+ ML teams that have cut their training costs by more than 50% with Lightly products.

Book a Demo

Free Download: Computer Vision Architecture Decision Tree

Picking DINOv3 or YOLO11 is easy. Getting it to run in production isn’t.

Learn how to do it properly. 👇

Thanks for submitting the form.