Skip to main content

New AI Models

Every notable new AI model from the major labs, newest first and curated from official announcements. Filter by provider and model type to see who shipped what, and when.

Provider
Type

698 releasesPage 4 of 35

ProviderModelReleasedLink
Hunyuan

hy3

Hy3 is a reasoning and agent model from Tencent Hunyuan, built on a MoE architecture with 295B total parameters, 21B activated per token, and a 256K context window. Tencent positions it as rivaling flagship models with 2-5x its parameters, with particular strength in software development, office productivity, and financial modeling. The model is open-sourced under Apache 2.0 and available through Tencent Cloud API.

July 3, 2026
Qwen

Wan2.7-Reference-To-Video-2026-06-12

Wan2.7-Reference-To-Video is the reference-to-video model in Alibaba's Tongyi Wanxiang 2.7 series. It accepts up to 5 mixed image and video references plus an optional voice reference, and outputs 720P or 1080P clips of 2 to 15 seconds with more stable character, prop, and scene consistency. It is suited to single-character performance and multi-character interaction. This 2026-06-12 snapshot is pinned for stable production use.

July 1, 2026
Qwen

Wan2.7-Text-To-Video-2026-06-12

Wan2.7-Text-To-Video-2026-06-12 is the June 12, 2026 snapshot of Alibaba's Wan2.7 text-to-video model, a fixed version suited to stable production use. It generates fluid video from text prompts with upgraded acting: nuanced emotional scenes, hard-hitting action, and more dramatic, well-paced shot transitions. Output covers 720P or 1080P at 2 to 15 seconds, with audio uploaded or auto-generated to match the visuals.

July 1, 2026
Anthropic

Claude Sonnet 5

Released on June 30, 2026, Claude Sonnet 5 is Anthropic's balanced-tier model and its most agentic Sonnet yet, pairing top-tier intelligence with coding and everyday professional work. It improves on Sonnet 4.6 across reasoning, tool use, coding, and knowledge work, and approaches Opus 4.8 performance at lower cost on agentic evals like BrowseComp and OSWorld-Verified. It suits multi-step software engineering, autonomous computer use, and business automation.

June 30, 2026
Google

Gemini Omni Flash

Gemini Omni Flash is Google's high-performance multimodal model for fast video generation and conversational video editing, introduced in public preview on June 30, 2026. It natively processes text, image, audio, and video, and uses the Interactions API to turn text prompts or still images into videos of 3 to 10 seconds with audio. Results can be refined through natural language conversation, with support for video extension, resolution upscaling, and interpolation.

June 30, 2026
Google

Nano Banana 2 Lite

Nano Banana 2 Lite (model code gemini-3.1-flash-lite-image) is the efficiency tier of Google's image generation family, released to GA on June 30, 2026. It targets sub-2 second end-to-end latency and significantly reduced compute costs, making it suited to high-volume interactive and real-time consumer applications. The model supports interleaved text and image generation, is optimized for 1K (1024x1024px) output, and handles fast multi-turn local edits such as color swaps, stickers, and background adjustments, with SynthID watermarking always on.

June 30, 2026
Grok

grok-4.5

Grok 4.5 is xAI's smartest model to date, built to excel at coding, agentic tasks, and knowledge work. It is served at fast-model speeds of 80 TPS, delivers roughly 2x the token efficiency of comparable leading models, and is priced at $2 per million input tokens and $6 per million output tokens. It is available today in Grok Build, in Cursor on all plans, and through the xAI API, and handles Office work in Excel, PowerPoint, and Word with similar care.

June 29, 2026
Kling

Kling 3.0

Kling 3.0 is the standard video generation model from Kuaishou's Kling AI, built on a unified multimodal training framework that consolidates text-to-video, image-to-video and reference-based generation into one model. It generates clips up to 15 seconds with flexible duration, adds Multi-Shot storytelling and Element consistency control, and outputs Native Audio in Chinese, English, Japanese, Korean and Spanish, including dialects and accents, alongside improved prompt adherence and image realism. It suits multi-shot cinematic storytelling, multi-character dialogue scenes and e-commerce advertising.

June 29, 2026
Kling

Kling 3.0 Omni

Kling 3.0 Omni is the all-in-one video generation model in Kuaishou's Kling 3.0 family, upgrading Kling VIDEO O1 with a deeply unified multimodal training framework for native audio and video output. It supports Native Audio and Multi-Shot storytelling, generates 3 to 15 second videos with shot-level control over duration, framing, perspective, and camera movement, and its Element system accepts video and voice references to keep characters consistent in both look and sound, fitting narrative work such as ads and short dramas.

June 29, 2026
Kling

Kling Image O1

Kling Image O1 is a reasoning image generation model that draws on a broad knowledge base and multimodal reasoning to interpret complex creative intent. It supports up to 10 reference images, keeping subject contours, core elements, and tonal qualities consistent across a series. Text instructions handle detail edits and style transfer while preserving original lighting and texture. Internal evaluations showed higher win rates than Nano Banana, Flux 1.0 Kontext Pro, and Seedream 4.0 on instruction-transformation and multi-image reference tasks, suited to IP character design and brand visuals.

June 29, 2026
OpenAI

GPT-5.6 Sol

GPT-5.6 Sol is OpenAI's flagship model, launched in limited preview in June 2026 alongside the balanced Terra and fast, affordable Luna tiers. It adds a new max reasoning effort and an ultra mode that leverages subagents, sets a new state of the art on Terminal-Bench 2.1, and outperforms GPT-5.5 on GeneBench v1 with fewer tokens while matching Mythos Preview on ExploitBench with roughly a third of the output tokens. Preview access runs via the API and Codex for select partners, priced at $5 input and $30 output per 1M tokens.

June 26, 2026
Qwen

Qwen-Image-2.0-Pro-2026-06-22

Qwen-Image-2.0-Pro-2026-06-22 is the latest dated snapshot of the full-strength Qwen-Image-2.0 Pro image model, released on Alibaba Cloud Model Studio on June 25, 2026. It unifies image generation and image editing in one model. Compared with the April snapshot, it renders text more accurately, accepts prompts up to 1k tokens, and delivers finer photorealistic detail with stronger semantic adherence. The pinned version number suits production workflows that need reproducible outputs.

June 25, 2026
ByteDance

Seed2.1

Seed2.1 is ByteDance's next-generation agentic model for real-world productivity, released in two sizes, Pro and Turbo. It brings major improvements to general agents and code engineering, covering end-to-end requirement understanding, coding, debugging, and validation in real development workflows. It also strengthens visual understanding, spatial reasoning, and long-context processing, and achieves SOTA results across multiple video understanding benchmarks.

June 23, 2026
OpenAI

GPT-5.5 Instant

GPT-5.5 Instant is OpenAI's updated default model for ChatGPT, rolling out to all users and available in the API as chat-latest. In internal evaluations it produced 52.5% fewer hallucinated claims than GPT-5.3 Instant on high-stakes prompts in medicine, law, and finance, with gains across visual reasoning, math, and science. Answers are tighter and less overformatted, and the model makes better use of context from past chats, files, and connected Gmail to personalize responses.

June 18, 2026
Qwen

Fun-ASR-Flash-2026-06-15

Fun-ASR-Flash-2026-06-15 is the dated June 2026 snapshot of the Bailing large-model ASR line on Alibaba Cloud Model Studio, a fit for production workloads that need consistent output. It transcribes audio up to 5 minutes across 30 languages, supports all seven major Chinese dialect groups, and refines punctuation prediction and text normalization so numbers, dates, and amounts follow standard written formats. Recognition of classical Chinese poetry is specifically optimized, and usage is billed per second of audio from USD 0.00003.

June 17, 2026
Zhipu AI

GLM-5.2

GLM-5.2 is Zhipu AI's flagship model for long-horizon tasks, open-sourced under an MIT license with a solid 1M token context window. An IndexShare sparse-attention design trims per-token FLOPs by 2.9x at full context, and coding strength takes a major step up over GLM-5.1: 81.0 on Terminal-Bench 2.1 and 62.1 on SWE-bench Pro, the best among open-source models and second only to Claude Opus 4.8 on FrontierSWE and PostTrainBench. Adjustable thinking effort levels balance performance against latency for coding agents and long-running engineering work.

June 17, 2026
Qwen

HappyHorse-1.1-I2V

HappyHorse-1.1-I2V is an image-to-video model from Alibaba Cloud Bailian that turns a single first-frame image into a 3 to 15 second clip with audio, offered at 480P, 720P, and 1080P with 24fps MP4 output. It interprets the input image more precisely and carries the creative intent forward, improving skin texture, cross-shot ID consistency, motion smoothness, text rendering stability, and audio-visual sync to produce detailed, coherent video.

June 16, 2026
Qwen

HappyHorse-1.1-R2V

HappyHorse-1.1-R2V is a reference-to-video model on Alibaba Cloud Model Studio that generates video from up to 9 reference images combined with a text prompt. Building on the earlier release, it delivers more stable consistency of subjects, scene style, and overall visuals, with stronger controllability across characters, scenes, and cinematography. Generated videos run 3 to 15 seconds at 480P, 720P, or 1080P resolution.

June 16, 2026
Qwen

HappyHorse-1.1-T2V

HappyHorse-1.1-T2V is a text-to-video generation model from Alibaba Cloud, offered on the Model Studio (Bailian) platform. It improves text comprehension, camera control, and motion generation over the previous release, translating creative intent more accurately into video with smoother motion, richer detail, and higher consistency in character action, scene atmosphere, and physics. It supports 3 to 15 second clips at 480P, 720P, or 1080P, targeting scenarios such as short dramas, e-commerce ads, brand marketing, and game CG.

June 16, 2026
Moonshot AI

Kimi K2.7 Code

Kimi K2.7 Code is a coding-focused agentic model from Moonshot AI, built on Kimi K2.6 with a Mixture-of-Experts architecture of 1T total and 32B activated parameters and a 256K context window. It targets long-horizon software engineering tasks, improving end-to-end completion while cutting thinking-token usage by about 30% versus K2.6, and scores higher on benchmarks such as Kimi Code Bench v2 and MCPMark-Verified. Weights are open under the Modified MIT License, deployable via vLLM and SGLang with OpenAI- and Anthropic-compatible APIs.

June 12, 2026

How to track new AI models

Four ways to read the board. Each one turns the stream of new AI models into a clear answer. Pick the view that fits your question.

01

Read the newest launches first

The board lists new AI models newest first, so the top row is always the latest launch from any lab. Every row links to the official announcement, so one click gets you to the source. Dates reflect the initial public release.

02

Follow one provider

Pick a lab in the Provider filter. OpenAI, Anthropic, Google, Meta, DeepSeek, Qwen and 20+ more each get their own view. Use it to read a lab's release rhythm: how often it ships, and which model types it bets on.

03

Compare model types

Filter by type: text/LLM, multimodal, image, video, or audio. The mix shows where the industry is pushing: text models still lead the count, while image and video generators ship in waves. Watch the balance shift month by month.

04

Catch up on what you missed

Page back through the timeline. Some months see dozens of launches, and the archive reaches back years. The board is a full release history, not just a feed of today's new AI models. Reconstruct any lab's year in a few scrolls.

New AI Models FAQ

What the board tracks, where the data comes from, and how to read it.

The board sorts releases newest first, so the top row is always the most recent launch from any provider. Recent months brought new AI models from labs like OpenAI, Anthropic, Google, Qwen and Zhipu AI. Every row links to the official announcement with its exact release date.

Twenty-nine providers and counting: OpenAI, Anthropic, Google, Meta, Microsoft, Amazon, NVIDIA, Mistral AI, DeepSeek, Qwen, Moonshot AI, Zhipu AI, ByteDance, xAI and more. Western and Chinese labs share one board, so a launch abroad still shows up the same day.

Five kinds: text/LLM, multimodal, image generation, video generation, and audio/speech. Text models carry the largest share of launches, but image and video generators ship in fast waves. The Type filter isolates any one stream.

Constantly. Major labs now ship new AI models on a weekly cadence, and busy weeks bring more than one frontier launch. Since 2024 the pace has kept climbing, and the board adds each release as its official announcement goes live.

Each row uses the date of the initial public release: the day the lab announced the model or opened it to users. Betas, rumors and paper preprints do not count. Every entry links to the official announcement, so you can verify the date yourself.

Yes. The board doubles as an archive reaching back to 2017. Page back to see how the release cadence exploded, from a handful of new AI models a year to hundreds. It is a searchable release history, not just a live feed.

Yes. The Provider filter narrows the board to one lab or any set of labs. Combine it with the Type filter to answer narrow questions, like every video model a single lab has shipped.

Yes, completely. Every release, filter and official link is free to browse, no signup needed. The board updates as new models are announced, so it pays to check back after every launch event.

Curated from official provider announcements; dates reflect the initial public release.