OpenAI outlines safety risks and safeguards for long-horizon AI models
OpenAI describes new safety risks tied to long-horizon AI models in a newly published report.
OpenAI describes new safety risks tied to long-horizon AI models in a newly published report.
Moonshot AI released Kimi K3, a 2.8 trillion parameter Mixture-of-Experts model slated for open-weights release on July 27, 2026.
OpenAI discussed releasing a GPT-3-level language model capable of running on consumer hardware in 2022.
Claude Code v2.1.181 and later versions embed a Rust-based port of Bun.
Moonshot AI released Kimi K3, an open source model it says achieves frontier-level performance on internal evaluations.
OpenAI's CFO introduced a scorecard framework to assess AI return on investment.
Moonshot AI launched Kimi K3 as a frontier-class open-weights model with 2.8 trillion parameters and a 1 million-token context window.
Moonshot AI’s Kimi K3 is expected to close the performance gap with Anthropic’s Opus 4.8.
Thinky introduced Inkling, its first open-weights foundation model family, with a 975B total/41B active parameter Mixture-of-Experts architecture supporting text, image, and audio inputs and up to 1M-token context.
Thinking Machines Labs released Inkling, its first proprietary AI model, as an open-weight system with 975 billion total parameters and a 41 billion active parameter design for efficiency.
Bilibili released Index-1.9B, a series of four open small language models with 1.9B non-embedding parameters each.
OpenAI’s GPT-5.6 family introduces three variants—Luna, Terra, and Sol—with input/output pricing ranging from $1/$6 to $5/$30 per 1M tokens.
OpenAI unveiled GPT-5.6, a new family of models with three variants: Sol (workhorse), Terra (intermediate), and Luna (budget).
OpenAI's documentation states cloud Work conversations are not visible in the desktop app at launch.
Deutsche Telekom and OpenAI announced a multi-year partnership to integrate AI across Deutsche Telekom's operations.
OpenAI describes a new agentic system called ChatGPT Work designed to persist across multi-hour projects and execute actions across users' apps and files.
Gemma 4 introduces a new generation of open-weight, natively multimodal language models with dense and Mixture-of-Experts architectures ranging from 2.3B to 31B parameters.
Tencent’s Hy3 is a 295B-parameter Mixture-of-Experts model with 21B active parameters and 3.8B MTP layer parameters.
Newer Anthropic models (Opus 4.8, Sonnet 5) sometimes add invented keys to the edits[] array when calling third-party edit tools, causing schema validation failures.
Claude Sonnet 5 is positioned by Anthropic as having performance close to Opus 4.8 at lower prices.
Claude Sonnet 5 is Anthropic’s latest mid-size model, offering stronger agentic capabilities than its predecessor.
OpenAI engineers diagnosed rare infrastructure crashes using large-scale core dump analysis.
Ornith-1.0 is an open-weight model family released under the MIT license by DeepReinforce, the first model release from the organization.
Open model ecosystem is diversifying beyond a handful of dominant players, with contributions from niche companies globally.
OpenAI started a limited preview of the GPT‑5.6 series, including flagship Sol, balanced Terra, and low-cost Luna.
Gemini 3.5 Flash now includes built-in computer-use tooling for agentic tasks across platforms.
OpenAI announced support for a new global standards effort aimed at guiding the development of advanced AI systems.
Researchers describe "role confusion" as a core failure mode where models prioritize text style over content, enabling prompt injection attacks.
A 0.2B-parameter image inpainting model (Moebius) was converted from PyTorch to ONNX and run in a browser using WebGPU.
Z.ai released GLM-5.2 on June 13, 2026, to GLM Coding Plan members, with public weights and a blog post on June 16, 2026.
Subquadratic, a Miami-based AI startup, claims its sparse-attention LLM SubQ matches top coding performance while using far less compute.
Chinese lab Z.ai released GLM-5.2 as open weights under an MIT license on June 16, 2026, following a subscriber preview on June 13.
Two new Mixture-of-Experts models, DeepSeek-V4-Pro and DeepSeek-V4-Flash, support one-million-token contexts with improved efficiency.
OpenAI introduced spend controls and usage analytics for ChatGPT Enterprise to help organizations manage costs.
AMIE, Google DeepMind’s research AI system for medical reasoning, matched primary care physicians in overall disease management reasoning in a blinded study with patient actors.
MolmoMotion predicts future 3D trajectories of object points from video frames and language instructions, outperforming prior motion forecasting methods.
GLM-5.2 is an open-source model optimized for long-horizon coding tasks with a solid 1M-token context.
OpenAI introduced Deployment Simulation, a method to predict AI model behavior before deployment.
Stability AI released Stable Audio 3.0, a family of four models ranging from 459M to 2.7B parameters, with top models generating music up to 6 minutes 20 seconds long—double the length of its 2024 predecessor.
Allen Institute released OlmoEarth v1.1, an update to its transformer-based remote sensing model designed for satellite imagery analysis.
Google reported that AI Mode, its conversational search feature introduced one year ago in the U.S., has reached over one billion monthly active users globally and queries have doubled every quarter since launch.
OpenAI launched GPT-Realtime-2, positioning it as a voice-to-speech model with 'GPT-5-class reasoning' capable of tool use, interruption recovery, and longer conversations via expanded 128K context window.
OpenAI announced GPT-5.5 and GPT-5.5-Cyber as part of its Trusted Access for Cyber initiative
Zyphra published a technical report on ZAYA1-8B, a reasoning-focused mixture-of-experts model with 700M active parameters and 8B total parameters trained entirely on AMD compute infrastructure.
OpenAI has published new realtime voice models for developers to use through its API.
Anthropic released version 0.100.0 of its Python SDK on May 6, 2026. The release adds support for managed agents, multiagents, outcomes, webhooks, and vault validation. Bug fixes include adjustments to webhook configuration.
OpenAI released v2.34.0 of its official Python SDK on May 4, 2026, adding support for per-endpoint Admin API keys and bearer-free admin authentication.
OpenAI is scaling Stargate, its compute infrastructure initiative, to expand data center capacity
Google describes TPUs as custom chips designed specifically to run AI models, emphasizing their ability to perform complex mathematical operations at scale.
OpenAI announced its GPT models, Codex, and Managed Agents are now available on AWS infrastructure.