AI news & updates
361 events
- Model launchMajorQwen releases Qwen3.6 Plus
Qwen 3.6 Plus builds on a hybrid architecture that combines efficient linear attention with sparse mixture-of-experts routing, enabling strong scalability and high-performance inference. Compared to the 3.5 series, it de
- Model launchZ.ai releases GLM 5V Turbo
GLM-5V-Turbo is Z.ai’s first native multimodal agent foundation model, built for vision-based coding and agent-driven tasks. It natively handles image, video, and text inputs, excels at long-horizon planning, complex cod
- Model launchMajorSpaceXAI releases Grok 4.20 Multi-Agent
Grok 4.20 Multi-Agent is a variant of SpaceXAI’s Grok 4.20 designed for collaborative, agent-based workflows. Multiple agents operate in parallel to conduct deep research, coordinate tool use, and synthesize information.
- Model launchMajorAlibaba releases Qwen 3.6 Plus
- Model launchMajorAlibaba releases Qwen 3.6 Plus (2026-04-02)
- Model launchMajorGoogle releases Lyria 3 Pro Preview
Full-length songs are priced at $0.08 per song. Lyria 3 is Google's family of music generation models, available through the Gemini API. With Lyria 3, you can generate high-quality, 48kHz...
- Model launchMajorGoogle releases Lyria 3 Clip Preview
30 second duration clips are priced at $0.04 per clip. Lyria 3 is Google's family of music generation models, available through the Gemini API. With Lyria 3, you can generate...
- Model launchMajorOpenAI releases GPT-5.4 Nano
GPT-5.4 nano is the most lightweight and cost-efficient variant of the GPT-5.4 family, optimized for speed-critical and high-volume tasks. It supports text and image inputs and is designed for low-latency...
- Model launchMajorOpenAI releases GPT-5.4 Mini
GPT-5.4 mini brings the core capabilities of GPT-5.4 to a faster, more efficient model optimized for high-throughput workloads. It supports text and image inputs with strong performance across reasoning, coding,...
- Model launchMajorOpenAI releases GPT-5.4 Nano (batch)
GPT-5.4 nano is the most lightweight and cost-efficient variant of the GPT-5.4 family, optimized for speed-critical and high-volume tasks. It supports text and image inputs and is designed for low-latency...
- Model launchMajorOpenAI releases GPT-5.4 Mini (batch)
GPT-5.4 mini brings the core capabilities of GPT-5.4 to a faster, more efficient model optimized for high-throughput workloads. It supports text and image inputs with strong performance across reasoning, coding,...
- Model launchMajorOpenAI releases GPT-5.4 mini (xhigh)
- Model launchMajorOpenAI releases GPT-5.4 mini (none)
- Model launchMajorOpenAI releases GPT-5.4 nano (no thinking)
- Model launchMajorOpenAI releases GPT-5.4 nano (low)
- Model launchMajorOpenAI releases GPT-5.4 mini (low)
- Model launchMajorOpenAI releases GPT-5.4 mini (medium)
- Model launchMajorOpenAI releases GPT-5.4 nano (medium)
- Model launchMajorOpenAI releases GPT-5.4 nano (xhigh)
- Model launchMajorOpenAI releases GPT-5.4 mini (unknown thinking)
- Model launchZ.ai releases GLM 5 Turbo
GLM-5 Turbo is a new model from Z.ai designed for fast inference and strong performance in agent-driven environments such as OpenClaw scenarios. It is deeply optimized for real-world agent workflows...
- Model launchByteDance Seed releases Seed-2.0-Lite
Seed-2.0-Lite is a versatile, cost‑efficient enterprise workhorse that delivers strong multimodal and agent capabilities while offering noticeably lower latency, making it a practical default choice for most production w
- Model launchMajorOpenAI releases GPT-5.4
GPT-5.4 is OpenAI’s latest frontier model, unifying the Codex and GPT lines into a single system. It features a 1M+ token context window (922K input, 128K output) with support for...
- Model launchMajorOpenAI releases GPT-5.4 Pro
GPT-5.4 Pro is OpenAI's most advanced model, building on GPT-5.4's unified architecture with enhanced reasoning capabilities for complex, high-stakes tasks. It features a 1M+ token context window (922K input, 128K...
- Model launchMajorOpenAI releases GPT-5.4 (batch)
GPT-5.4 is OpenAI’s latest frontier model, unifying the Codex and GPT lines into a single system. It features a 1M+ token context window (922K input, 128K output) with support for...
- Model launchMajorOpenAI releases GPT-5.4 Pro (batch)
GPT-5.4 Pro is OpenAI's most advanced model, building on GPT-5.4's unified architecture with enhanced reasoning capabilities for complex, high-stakes tasks. It features a 1M+ token context window (922K input, 128K...
- Model launchMajorOpenAI releases GPT-5.4 (medium)
- Model launchMajorOpenAI releases GPT-5.4 (none)
- Model launchMajorOpenAI releases GPT-5.4 (low)
- Model launchMajorOpenAI releases GPT-5.4 Pro (xhigh)
- Model launchMajorOpenAI releases GPT-5.4 (xhigh)
- Model launchMajorOpenAI releases GPT-5.4 Pro (web)
- Model launchMajorOpenAI releases GPT-5.4 (unknown thinking)
- Model launchMajorOpenAI releases GPT-5.4 Pro (no thinking)
- Model launchInception releases Mercury 2
Mercury 2 is an extremely fast reasoning LLM, and the first reasoning diffusion LLM (dLLM). Instead of generating tokens sequentially, Mercury 2 produces and refines multiple tokens in parallel, achieving...
- Model launchMajorGoogle releases Gemini 3.1 Flash Lite
Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, image, video, audio, and PDF inputs, and is designed for lightweight agentic...
- Model launchMajorGoogle releases Gemini 3.1 Flash Lite Preview
Gemini 3.1 Flash Lite Preview is Google's high-efficiency model optimized for high-volume use cases. It outperforms Gemini 2.5 Flash Lite on overall quality and approaches Gemini 2.5 Flash performance across...
- Model launchByteDance Seed releases Seed-2.0-Mini
Seed-2.0-mini targets latency-sensitive, high-concurrency, and cost-sensitive scenarios, emphasizing fast response and flexible inference deployment. It delivers performance comparable to ByteDance-Seed-1.6, supports 256
- Model launchMajorGoogle releases Nano Banana 2 (Gemini 3.1 Flash Image Preview)
Gemini 3.1 Flash Image Preview, a.k.a. "Nano Banana 2," is Google’s latest state of the art image generation and editing model, delivering Pro-level visual quality at Flash speed. It combines...
- Model launchMajorQwen releases Qwen3.5-Flash
The Qwen3.5 native vision-language Flash models are built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. Compared to t