Skip to content
AI Primer
TOPIC5 stories

Computer Vision

Stories, products, and related signals connected to this tag in Explore.

RELEASE1st September
World Labs releases Atlas multimodal world model

World Labs says Atlas combines visual generation with scene reconstruction. It can reconstruct scenes from images, generate camera-controlled frames, and reframe video.

NEWS31st August
Benchmark author says GLM-5.3-Flash rerun improves scores after routing error

A benchmark author says OpenRouter likely routed GLM-5.3-Flash requests to quantized endpoints because precision was not pinned. Twelve reruns using pinned FP8 and self-hosted inference improved results, suggesting earlier scores may have reflected routing.

RELEASE1w ago
Perceptron releases Isaac 0.5 open weights for robot control

Perceptron released weights, inference code, and training details for Isaac 0.5, an embodied model for video perception, reasoning, and robot control. The 36B dynamic-MoE model was trained on 1 million hours of video.

RELEASE1mo ago
OpenBMB releases MiniCPM-Robot models and PhyAI runtime with 33-36 Hz throughput claim

OpenBMB open-sourced MiniCPM-RobotManip, MiniCPM-RobotTrack and PhyAI, claiming local robot tracking, robot memory and throughput gains from 10 Hz to 33-36 Hz. The release packages model artifacts and a runtime path for local robot perception and manipulation experiments.

RELEASE1mo ago
Meta launches Muse Image with reasoning, search, code, and self-refinement

Meta launched Muse Image in its apps and previewed Muse Video from the same media-generation family. Meta says Muse Image can reason, search, write code, self-refine, and use test-time compute before generating images.

AI PrimerAI Primer

Your daily guide to AI tools, workflows, and creative inspiration.

© 2026 AI Primer. All rights reserved.