Composer 2.5 is now available inside Grok Build.
Composer 2.5 is a fast, highly intelligent model that excels on long-running tasks and following complex instructions.
Introducing Qwen3.7-Plus — a multimodal agent model that unifies vision and language into one versatile agent foundation.
Multimodal interactive hybrid agent: unified GUI & CLI operation across visual and text tasks
Versatile coding agent & productivity assistant with full-modality input
Visual Agent: perception, reasoning, grounding, and search-augmented QA
Cross-harness generalization across diverse agent frameworks
One model. Sees, thinks, codes, acts.
Now available via API on Alibaba Cloud Model Studio. Try it — let us know what you build.
Blog:
https://
qwen.ai/blog?id=qwen3.
7-plus
…
Qwen Studio:
https://
chat.qwen.ai/?models=qwen3.
7-plus
…
API:
https://
modelstudio.console.alibabacloud.com/ap-southeast-1
?tab=doc#/doc/?type=model&url=2840914_2&modelId=qwen3.7-plus&serviceSite=international
…
Sarvam AI has open-sourced two powerful reasoning models — Sarvam 30B and Sarvam 105B — trained from scratch with all data, model research, and inference optimization done in-house. The models 'punch above their weight' in global benchmarks while excelling in Indian languages. The 30B model uses classic Grouped Query Attention (GQA) while the 105B uses a different architecture approach.
Anthropic partnered with Mozilla to test Claude's vulnerability research capabilities. Opus 4.6 found 22 vulnerabilities in Firefox, 14 high-severity, representing a fifth of all high-severity bugs Mozilla remediated in 2025. 2.9M views.
A statement from Anthropic CEO Dario Amodei addressing the company's position on discussions with the Department of War. The post garnered massive engagement with 2.3M views and 42K likes.
Anthropic partnered with Mozilla to test Claude's ability to find security vulnerabilities in Firefox's source code. Opus 4.6 scanned nearly 6,000 C++ files, submitted 112 reports, and confirmed 22 vulnerabilities — 14 rated high-severity by Mozilla, representing roughly one-fifth of all high-severity Firefox bugs remediated in 2025. This demonstrates a major breakthrough in AI-assisted security auditing.
Anthropic CEO Dario Amodei published a statement titled 'Where things stand with the Department of War' on Anthropic's website, amid growing controversy about AI companies' involvement with defense and military applications. The statement garnered significant attention with 5,000 likes and 2.3 million views.
Codex Security is an AI application security agent that analyzes project context to detect, validate, and patch complex vulnerabilities with higher confidence and less noise.
Descript uses OpenAI models to scale multilingual video dubbing, optimizing translations for both meaning and timing so dubbed speech sounds natural across languages.
Anthropic released a study examining which jobs AI can theoretically replace versus which ones it's actually automating. Computer & math roles show 94% theoretical exposure, legal ~90%, and management, architecture, arts & media all 60%+. However, observed real-world usage is only a fraction of theoretical capability — though the gap is closing fast.
Dwarkesh Patel interviews historian Ada Palmer about Gutenberg, the printing press, Renaissance Florence, and the parallels between historical technological revolutions and AI.
Cognitive Revolution episode with Dan Balsam & Tom McGrath from Goodfire on using interpretability to reduce hallucination, discover Alzheimer's biomarkers, and separate memorization from reasoning.
OpenAI publishes evaluation suite and research paper on Chain-of-Thought Controllability. GPT-5.4 Thinking shows low ability to obscure its reasoning, suggesting CoT monitoring remains a useful safety tool.
OpenAI launches GPT-5.4, their most factual and efficient model. Brings advances in reasoning, coding, and agentic workflows into one frontier model. Available in ChatGPT, API, and Codex. 6.3M views, 23K likes.
OpenAI introduces CoT-Control and finds reasoning models struggle to control their chains of thought, reinforcing monitorability as an AI safety safeguard.
Introducing GPT-5.4, OpenAI’s most most capable and efficient frontier model for professional work, with state-of-the-art coding, computer use, tool search, and 1M-token context.
OpenAI introduces ChatGPT for Excel and new financial app integrations, powered by GPT-5.4 to accelerate modeling, research, and analysis in regulated environments.
A new preprint extends single-minus amplitudes to gravitons, with GPT-5.2 Pro helping derive and verify nonzero graviton tree amplitudes in quantum gravity.
Latent Space talk with Alex Atallah about how the first LLM aggregator got started, weird early growth moments, architecture challenges, and future plans.
Axios COO Allison Murphy explains how the company uses AI to support local reporters, streamline newsroom workflows, and deliver high-impact local journalism at scale.
Gemini 3.1 Flash-Lite outperforms 2.5 Flash with faster performance at lower price. Features new 'thinking levels' to dial in reasoning for different tasks. Available in preview via Gemini API in Google AI Studio. 1.7M views, 8.9K likes.
Cognitive Revolution episode with UK AISI Chief Scientist Geoffrey Irving surveying the AI landscape, discussing how jailbreaking frontier models is getting harder but the AISI Red Team has never failed.
Details on OpenAI’s contract with the Department of War, outlining safety red lines, legal protections, and how AI systems will be deployed in classified environments.
Anthropic releases an official statement addressing comments made by Secretary of War Pete Hegseth regarding AI and national security. The post went viral with 17M views and 56K likes.
Today we’re announcing $110B in new investment at a $730B pre money valuation. This includes $30B from SoftBank, $30B from NVIDIA, and $50B from Amazon.
Microsoft and OpenAI continue to work closely across research, engineering, and product development, building on years of deep collaboration and shared success.
Stateful Runtime for Agents in Amazon Bedrock brings persistent orchestration, memory, and secure execution to multi-step AI workflows powered by OpenAI.