Alibaba’s Tongyi Lab Releases Qwen-Audio-3.0-TTS, Text-to-Speech Model with Flash and Plus Tiers in All 16 Languages

Alibaba’s Tongyi Lab Releases Qwen-Audio-3.0-TTS, Text-to-Speech Model with Flash and Plus Tiers in All 16 Languages

Alibaba’s Tongyi Lab has been released Qwen-Audio-3.0-TTSa production-oriented text-to-speech (TTS) system. The model comes in two versions from the same list. Flash direct real-time interaction. in addition aimed at high quality workmanship. Both are delivered as managed models via Alibaba Cloud Model Studio, not as downloadable weights. The release focuses on four things that engineers … Read more

Someone’s MiniCPM5-1B Fine-tuned for OpenBMB in Claude Fable 5 Tracks for Posting a Local Thinking Model 657MB

Someone’s MiniCPM5-1B Fine-tuned for OpenBMB in Claude Fable 5 Tracks for Posting a Local Thinking Model 657MB

Community developer, GnLOLot, has published a fully functional 1B model for local hardware. Model i MiniCPM5-1B-Claude-Opus-Fable5-Thinkingwith GGUF build runtimes compatible with llama.cpp. It does not require an API key and does not make cloud calls. Proposed Model The model is built on openbmb/MiniCPM5-1B. That base is the original, documented release from OpenBMB. It is a … Read more

Feyn AI Releases SQRL, a Family of Text-to-SQL Models That Tests the Database Before Writing a Query

Feyn AI Releases SQRL, a Family of Text-to-SQL Models That Tests the Database Before Writing a Query

Most text-to-SQL systems treat the function as a translation. Feyn AI (a YC-backed startup) reframes it when it’s tested. Feyn’s team released SQRL, a family of models that turn natural language queries into SQL. Instead of generating a query on the fly, SQRL can scan the database first. This allows it to resolve ambiguity and … Read more

Alibaba Previews Qwen3.8-Max, 2.4 Trillion-Parameter Multimodal Model, Days After Moonshot’s Kimi K3 Open Weight Launch

Alibaba Previews Qwen3.8-Max, 2.4 Trillion-Parameter Multimodal Model, Days After Moonshot’s Kimi K3 Open Weight Launch

On July 19th, Alibaba’s Qwen team previewed Qwen3.8-Max-Preview, which is the flagship of the Qwen family. The research team describes it as a 2.4 trillion parameter model, ‘second only to Fable 5’ among the systems it estimates. The preview is live now. Benchmark table, model card, and license are not … Read more

Perplexity AI Releases WANDR: An Open Benchmark for Research Agents That Need to Search Wider and Deeper

Perplexity AI Releases WANDR: An Open Benchmark for Research Agents That Need to Search Wider and Deeper

Research agents are already handling the actual information work today. Teams submit competitive maps, due diligence, and literature reviews to them. However, most benchmarks evaluate a single response, not large evidence-based collections. Confusion is directed at the gap with a new open benchmark. Confusion is released WANDR (Wide AND Deep Research). It is an open … Read more

Kimi K3 vs DeepSeek V4 Pro vs GLM-5.2: Open Trillion-Scale MoE Models Compared to Benchmarks, Licensing, and Supply Costs

Kimi K3 vs DeepSeek V4 Pro vs GLM-5.2: Open Trillion-Scale MoE Models Compared to Benchmarks, Licensing, and Supply Costs

Three Chinese labs now hold the top of the open weight leaderboard. Moonshot AI’s Kimi K3, DeepSeek V4 Pro, and Zhipu AI’s GLM-5.2 are all outstanding Mixture-of-Experts (MoE) models with context windows of one million tokens. Each targets long-horizon coding and agent workloads. This article compares them on three axes that the AI ​​team decides … Read more

Fine-Tuning Qwen3 with LoRA Using NVIDIA NeMo AutoModel: Single-GPU Google Colab Workflow Complete Tutorial

Fine-Tuning Qwen3 with LoRA Using NVIDIA NeMo AutoModel: Single-GPU Google Colab Workflow Complete Tutorial

In this tutorial, we build end to end NVIDIA NeMo AutoModel workflow in Google Colab and use a single GPU to test the same architecture for configuration-driven training that scales on distributed multi-GPU environments. We verify the available CUDA hardware and precision support, install NeMo AutoModel directly into its repository, upload the official Qwen3-0.6B LoRA … Read more

NVIDIA Releases DeepStream 9.1: Brings Agentic AI to Vision AI With 13 Skills and Multi-View 3D Tracking

NVIDIA Releases DeepStream 9.1: Brings Agentic AI to Vision AI With 13 Skills and Multi-View 3D Tracking

NVIDIA recently released DeepStream 9.1. The update addresses an ongoing issue with video analytics. Tracking a single object from multiple cameras often requires manual camera calibration and complex calculations. DeepStream 9.1 addresses this with two additions: Multi-View 3D Tracking (MV3DT) and AutoMagicCalib (AMC). Both go as agent capabilities of coding agents. As a result, developers … Read more

NVIDIA AI Releases Nemotron 3 Embedded: 8B Benchmark Open Embedded Cluster Ranked #1 in RTEB

NVIDIA AI Releases Nemotron 3 Embedded: 8B Benchmark Open Embedded Cluster Ranked #1 in RTEB

Embedding models determine which episodes the agent has seen. NVIDIA released Nemotron Model 3 Embed working on that layer. It targets RAG for production scale, agent recovery, code recovery, and agent memory. What is Nemotron 3 Embed? The model cluster includes three open test areas. Nemotron-3-Embed-8B-BF16 accuracy-first option. Nemotron-3-Embed-1B-BF16 carries the same design into a … Read more

Moonshot AI Releases Kimi K3: 2.8 Trillion Parameter Open MoE Model With Kimi Delta Attention and 1M Content

Moonshot AI Releases Kimi K3: 2.8 Trillion Parameter Open MoE Model With Kimi Delta Attention and 1M Content

Moonshot AI has just been released For me K3. 2.8-trillion parameter model with native view and 1 million token context window. Moonshot calls it the world’s first open 3T-class model. What is Kimi K3? The Kimi K3 is a small Mixture-of-Experts (MoE) model built on two architectural revisions. Those are Kimi Delta Attention (KDA) and … Read more