LLMs

Latest breakthroughs in Large Language Models

Articles

The AI-Model Network: Concept, Current State, and Future Directions
Life After Benchmark Saturation: A Case Study of CORE-Bench
Detecting and Controlling Sycophancy with Cascading Linear Features
Project Auto-World: Towards Automated Benchmarking of Neural
The Hitchhiker's Guide to Agentic AI: From Foundations to Systems
Neuro-Symbolic Drive: Rule-Grounded Faithful Reasoning for
RIFT-Bench: Dynamic Red-teaming for Agentic AI Systems
Beyond Fixed Budgets: Characterizing the Inelasticity and
Measuring Curriculum Alignment across Topical Coverage,
Deontic Policies for Runtime Governance of Agentic AI Systems