Induction Labs Photon-1 Simulates Desktops, Plays Checkers, and Models Billiard Physics From One Pretraining Run
Most agents that learn from video need to know what action produced each frame. Induction Labs is arguing that this...
Most agents that learn from video need to know what action produced each frame. Induction Labs is arguing that this...
Check on YouTube
Black Forest Labs (BFL) has released FLUX 3, a multimodal foundation model that learns from images, videos and audio inside...
The KwaiKAT Team at Kuaishou has introduced the KAT-Coder-V2.5. It is a coding model trained to operate inside real, executable...
Check on YouTube
Sakana AI has released Fugu-Cyber (model ID is fugu-cyber-v1.0), a cybersecurity-specialized addition to its Fugu orchestration family. It is not...
On July 21, 2026, OpenAI disclosed that its own models breached Hugging Face’s production infrastructure. The models were not attacking...
Check on YouTube
Datalab has released Marker 2, a full rewrite of its open source document conversion pipeline. Marker converts PDF, image, PPTX,...
In this tutorial, we build a complete workflow for running Baidu’s Unlimited-OCR model on document images and multi-page PDFs. We...
Andrew Ng has announced OpenWorker, an open-source desktop agent that produces finished work rather than conversation. OpenWorker asks the user...
Open speech recognition stopped being a Whisper monoculture some time in the last twelve months. In March 2026 Cohere released...
Cursor has made Cursor Router generally available for Teams and Enterprise plans. The system is a classifier that inspects each...
Four open source projects dominate LLM fine-tuning today. Unsloth, Axolotl, TRL, and LLaMA-Factory all wrap the same underlying PyTorch and...
Poolside has released Laguna S 2.1, a 118B-parameter open-weight model built for agentic coding. It is a Mixture-of-Experts (MoE) model...
Meta has released Astryx, an open source design system that is fully customizable and built to be operated by both...
Alibaba’s Tongyi Lab has released Qwen-Audio-3.0-TTS, a production-oriented text-to-speech (TTS) system. The model ships in two variants from the same...
A single 24GB card is the practical floor for serious local inference. It is enough for genuinely capable models, and...