Claude Opus 5 is Anthropic’s new flagship model, promising near-frontier intelligence at a better cost. It is now the default for Claude Max and the strongest option on Claude Pro, with strong results on coding, automation, knowledge work, and scientific tasks. Read more
Google introduced Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber for building faster, cheaper AI agents at scale. 3.6 Flash cuts output token use versus 3.5 Flash while improving coding and multimodal work, Flash-Lite targets speed, and Flash Cyber pairs with CodeMender for security tasks. Read more
OpenAI is rolling out Health in ChatGPT for U.S. users 18 and older on web and iOS. Users can connect Apple Health and supported medical records so ChatGPT can compare results, summarize changes, and answer health questions with personal context, while OpenAI says connected health data is not used to train foundation models or target ads. Read more
OpenAI and Hugging Face disclosed a security incident triggered during an internal cyber-capability evaluation. OpenAI says evaluation models found a way out of the isolated test setup, reached Hugging Face systems while pursuing benchmark answers, and both companies are now investigating and tightening controls. Read more
NVIDIA and other large tech companies published a letter urging U.S. policymakers to avoid premature restrictions on open-weight AI models. They argue that open models help startups, researchers, and defenders inspect systems, compete, and build AI without lock-in. Anthropic is the only large AI lab that has not signed this letter. Read more
Do you prefer to watch instead of read? Check out this video covering the top AI news this week:
Black Forest Labs previewed FLUX 3, a coming multimodal model for image, video, audio, and action prediction. Details are still limited, but the pitch is a single creative model that can make realistic outputs across styles instead of separate tools for each media type. Read more
Alibaba launched Qwen-Image-3.0, its third-generation image generation model focused on useful, information-dense visuals. The official page was thin in direct fetches, but Qwen’s release materials describe stronger small-text rendering, richer layouts, realistic detail, and knowledge-heavy images such as documents and diagrams. Read more
Poolside released Laguna S 2.1, a 118B-parameter coding model with only 8B active parameters per token. It supports up to a 1M-token context and is designed for long-horizon software work, with Poolside highlighting strong coding benchmarks and full evaluation trajectories for every final trial. Read more
Microsoft’s Mage team introduced Mage-Flow, a compact 4B image generation and editing model family built for fast high-resolution work. Its Turbo versions generate 1024-square images in about 0.59 seconds and edits in about 1.02 seconds on a single A100, showing how smaller image models can still feel interactive. Read more
GLM is Z.AI’s flagship open-source model built for long, difficult engineering tasks like coding, debugging, and autonomous agent work. It delivers top-tier performance while being dramatically cheaper and faster than frontier closed models. Get 10% OFF HERE.
Nanbeige4.2-3B is a compact open-weight agent model designed for tool use, coding, reasoning, and local assistant workflows. Its Looped Transformer reuses layers to increase capacity without adding parameters, and the model card says it beats larger Qwen and Gemma models on several agent and code benchmarks. Read more
OpenAI’s docs now describe ChatGPT Voice for Chat, Work, and Codex in the ChatGPT desktop app. The feature lets users talk through ideas, start or steer tasks, check progress, and use voice dictation when they only need speech turned into prompt text. Read more
Google Quantum AI showed a reinforcement-learning approach that lets a quantum computer adjust while it runs. The system learns from error-correction signals to keep thousands of control settings stable, which matters because useful quantum programs may need to run for days or months without stopping. Read more
Stanford Medicine researchers used AI to find BRP, a natural peptide that may reduce appetite and body weight like semaglutide. In animal studies, BRP cut food intake and fat while avoiding clear signs of nausea, constipation, and major muscle loss, but it still needs human testing. Read more
NVIDIA Research introduced SANA-Video 2.0, an efficient video generation model built to make high-quality video on a single GPU. The project reports 720p five-second generation in about 13 seconds on one H100 for its optimized 5B pipeline, which could make video models cheaper to run and test. Read more
Typeless is an intelligent AI voice dictation tool designed to turn your speech into polished, well-structured text in real time. It automatically cleans up your speech by removing filler words like “um,” fixing mid-sentence corrections, and formatting lists across various apps and devices. Try it for free today!
ShotPlan is a research system for making multi-shot videos with precise scene changes and camera moves. It adds learnable planning tokens to a text-to-video diffusion model so creators can ask for hard cuts, cross-fades, and timed camera motion instead of hoping the model guesses the edit. Read more
Stanford-led researchers introduced masked visual actions as a way to control video world models across different robots and objects. Instead of feeding a robot-specific action format, the model uses a partial action video and predicts what happens next, which could help plan or evaluate robot behavior before real-world execution. Read more
Open Dreamer is an open effort to reproduce frontier-style world model training and share the lessons along the way. The team is open-sourcing the model and training code, starting with CoinRun as a smaller testbed before aiming at more complex Minecraft-like world modeling. Read more
HOMIE is a video personalization framework focused on keeping both people and objects consistent in generated videos. It targets cases where humans interact with objects, logos, or multi-view references, aiming to preserve identity, text, and object relationships better than earlier subject-driven video methods. Read more
Meet Fish Audio, the most expressive AI voice model. Core features include text-to-speech, instant voice cloning, and a library of over 2 million voices. Create natural, emotionally rich AI voices built for your workflows. Try it for free.

Comments
Nothing yet. Say the first thing.
Sign in to join the conversation.