Rethinking Toxic Data in LLM Pretraining: A Co-Design Approach for Improved Steerability and Detoxification
In the pretraining of LLMs, the quality of training data is crucial...
PwC Releases Executive Guide on Agentic AI: A Strategic Blueprint for Deploying Autonomous Multi-Agent...
In its latest executive guide, “Agentic AI – The New Frontier in...
Reinforcement Learning, Not Fine-Tuning: Nemotron-Tool-N1 Trains LLMs to Use Tools with Minimal Supervision and...
Equipping LLMs with external tools or functions has become popular, showing great...
Implementing an LLM Agent with Tool Access Using MCP-Use
MCP-Use is an open-source library that lets you connect any LLM to...
RL^V: Unifying Reasoning and Verification in Language Models through Value-Free Reinforcement Learning
LLMs have gained outstanding reasoning capabilities through reinforcement learning (RL) on correctness...
OpenAI Releases HealthBench: An Open-Source Benchmark for Measuring the Performance and Safety of Large...
OpenAI has released HealthBench, an open-source evaluation framework designed to measure the...
Multimodal AI Needs More Than Modality Support: Researchers Propose General-Level and General-Bench to Evaluate...
Artificial intelligence has grown beyond language-focused systems, evolving into models capable of...
A Step-by-Step Guide on Building, Customizing, and Publishing an AI-Focused Blogging Website with Lovable.dev...
In this tutorial, we will guide you step-by-step through creating and publishing...
PrimeIntellect Releases INTELLECT-2: A 32B Reasoning Model Trained via Distributed Asynchronous Reinforcement Learning
As language models scale in parameter count and reasoning complexity, traditional centralized...
AG-UI (Agent-User Interaction Protocol): An Open, Lightweight, Event-based Protocol that Standardizes How AI Agents Connect...
The current generation of AI agents has made significant progress in automating...






















