Tencent Hunyuan Hy4 Preview Launches 770B MoE Frontier Model
The August 28 2026 release delivers competitive engineering task performance at lower per token costs while integrating across productivity applications and third party platforms.
Releases, benchmarks and capability jumps from OpenAI, Anthropic, Google DeepMind, Meta and xAI.
Frontier models are the largest, most capable general-purpose AI systems at the edge of what is technically possible — today that means GPT-class models from OpenAI, Anthropic's Claude, Google DeepMind's Gemini, Meta's Llama and xAI's Grok. This section tracks every major release, benchmark result and capability shift across the leading labs, with sourced analysis of what each change means for the people who build on these models.
The August 28 2026 release delivers competitive engineering task performance at lower per token costs while integrating across productivity applications and third party platforms.
The integration supplies US and European developers with access to a 753B-parameter coding model and its 321B multimodal counterpart under zero data retention rules.
InclusionAI has released a finance-specialized variant of its Ling 3.0 Flash model that improves benchmark performance and offers free access periods before open-sourcing the weights.
The update equips developers with tools for incremental video building and resolution flexibility, advancing production workflows in a competitive AI media landscape.
Ant Group-backed InclusionAI deploys a finance-specialized model through API gateways on August 28, 2026, with open weights scheduled for the following week amid competition from other Chinese developers.
The launch provides open access to a high-efficiency natively multimodal model that narrows the performance gap with closed systems while cutting costs substantially.
ZhipuAI launches GLM-5.3-Flash as the first multimodal GLM-5 model with hybrid attention architecture, 1M context, and MIT-licensed weights, matching top-tier intelligence while running on domestic Chinese chips at sharply lower prices.
The August 26 release followed stealth testing on OpenRouter as ox-alpha, where the multimodal model topped weekly traffic before Z.ai disclosed its 320B-parameter design and MIT-licensed weights on Hugging Face.
Alibaba's open-weight release combines a sparse MoE design with architecture innovations that preview Qwen4, offering developers lower barriers to high-performance AI tools amid competition from closed models.
Open-weight MoE releases from Alibaba and Zhipu AI emphasize low active parameter counts and reduced costs while matching high-end coding and agentic benchmarks.
Independent evaluation places the xAI model alongside Anthropic's Claude Opus 5 at the leading score for real-world agentic tasks including tool use and autonomous problem-solving.
The release combines asynchronous Agent Team coordination, a locally deployable 35B model, and an open harness to enable executable long-horizon work in finance, research, and professional domains.
The August 14 2026 release maintains the 743B base from GLM-5.2 while scaling post-training to outperform Kimi K3 and DeepSeek V4 Pro on Terminal-Bench 3.0, CyberGym, and GDPval-AA v2 at prior pricing.
The anonymous model released August 20 delivers competitive long-horizon coding performance alongside DeepSeek V4 Pro while remaining free with a 1,048,576 token context window on OpenRouter.
The model appears on Nous Portal and Cline without official announcement, providing high capacity access to developers amid questions about its Zhipu origins.
The August 21, 2026, launch adds image input support to the API while preserving text performance and advancing multimodal agent benchmarks toward Opus-4.8 levels.
The 27B-parameter multimodal model delivers competitive coding and agentic benchmarks while supporting local deployment on consumer hardware under an open license.
The company now runs its frontier model in security scans for every Enterprise plan subscriber, delivering advanced vulnerability detection through a controlled proxy while direct model access stays limited to vetted partners.
Unsloth and orcarouter adapt the Qwen3.8-27B model into GGUF formats that retain native multimodal features and long context while delivering measurable accuracy gains through Dynamic V3.0 quantization.
The experimental model integrates image understanding into the V4-Flash framework without increasing costs or sacrificing text performance, narrowing the gap with leading multimodal systems on agent benchmarks.
A frontier model is a general-purpose AI system trained at the largest compute scales — models like GPT-class, Claude, Gemini, Llama and Grok — whose capabilities exceed prior systems on broad benchmarks.
Major labs ship a flagship roughly every 6–12 months, with point upgrades in between. This hub updates on every confirmed release.
Each article here cites the primary benchmark source — the lab model card or an independent eval — so you can verify the numbers directly.