#large language model
Large Language Model: 13 AI articles covering large language model news, analysis, and research
Articles
DeepSeek-V4.1-Flash Launches with 1M Context, FP4 KV Cache, and Cross-Layer Attention ReuseNEWβ9
DeepSeek-V4.1-Flash launches with 1M context, FP4 KV cache, and cross-layer attention reuseβcutting KV memory to 890 bytes per token for input-heavy agent workl...
OpenAI Releases GPT-6 Astra: A 1.05M-Context Computer-Use Model Gated Behind a βCriticalβ Cyber Thresholdβ8
OpenAI launches GPT-6 Astra, a 1.05M-context computer-use model gated behind a critical cyber threshold, sparking safety and deployment questions.
Meta AI Unveils Muse Spark 1.3: Agentic Coding Model Cuts Tool Calls by ~20% and Token Usage by ~25% vs. Predecessorβ7
Meta AI launches Muse Spark 1.3 agentic coding model with 20% fewer tool calls and 25% less token usage than 1.2, boosting efficiency for production workflows.
How to Run a Chatbot on Your Own Computerβ9
Learn how to run a chatbot on your own computer with this guide covering hardware requirements, model selection, and step-by-step installation for privacy-focus...
Z.ai Ships GLM-5.3 Without Retraining the Base Model: Better atβ8
Z.ai's GLM-5.3 boosts coding and cybersecurity via post-training, no retraining. Scores surge on complex benchmarks; open weights release in two weeks.
NVIDIA Unveils Nemotron 3.5 Lightning: A 30B Open MoE with 3Bβ8
NVIDIA launches Nemotron 3.5 Lightning, a 30B open MoE with 3B active params, plus NeMo Switchyard router for faster, efficient AI agents.
Claude Fable 5 AI Finds a Tiny Formula That Topples anβ9
Claude Fable 5 AI cracks an 87-year-old math conjecture, finding a counterexample to the Jacobian problem, marking a major AI-driven mathematical breakthrough.
Alibaba Qwen Unveils Qwen3.8-Max: A 2.4-Trillion-Parameter MoEβ6
Alibaba unveils Qwen3.8-Max, a 2.4T-parameter MoE model with open weights next week, plus Qwen3.8-27B for practical deployment.
DeepSeek Unveils DeepSeek-V4-Flash-0731: Major Agentic andβ8
DeepSeek launches V4-Flash-0731 with DSpark speculative decoding and Responses API, boosting agentic workflows and coding speed.
KwaiKAT Team Releases KAT-Coder-V2.5: An Agentic Coding Modelβ8
KwaiKAT Team launches KAT-Coder-V2.5, an agentic AI trained on 100K+ verifiable repositories, achieving state-of-the-art code generation and debugging with 72% ...
