Rethinking Toxic Data in LLM Pretraining: A Co-Design Approach for Improved Steerability and Detoxification

In the pretraining of LLMs, the quality of training data is crucial...

PwC Releases Executive Guide on Agentic AI: A Strategic Blueprint for Deploying Autonomous Multi-Agent...

In its latest executive guide, “Agentic AI – The New Frontier in...

Reinforcement Learning, Not Fine-Tuning: Nemotron-Tool-N1 Trains LLMs to Use Tools with Minimal Supervision and...

Equipping LLMs with external tools or functions has become popular, showing great...

Implementing an LLM Agent with Tool Access Using MCP-Use

MCP-Use is an open-source library that lets you connect any LLM to...

RL^V: Unifying Reasoning and Verification in Language Models through Value-Free Reinforcement Learning

LLMs have gained outstanding reasoning capabilities through reinforcement learning (RL) on correctness...

OpenAI Releases HealthBench: An Open-Source Benchmark for Measuring the Performance and Safety of Large...

OpenAI has released HealthBench, an open-source evaluation framework designed to measure the...

Multimodal AI Needs More Than Modality Support: Researchers Propose General-Level and General-Bench to Evaluate...

Artificial intelligence has grown beyond language-focused systems, evolving into models capable of...

A Step-by-Step Guide on Building, Customizing, and Publishing an AI-Focused Blogging Website with Lovable.dev...

In this tutorial, we will guide you step-by-step through creating and publishing...

PrimeIntellect Releases INTELLECT-2: A 32B Reasoning Model Trained via Distributed Asynchronous Reinforcement Learning

As language models scale in parameter count and reasoning complexity, traditional centralized...

AG-UI (Agent-User Interaction Protocol): An Open, Lightweight, Event-based Protocol that Standardizes How AI Agents Connect...

The current generation of AI agents has made significant progress in automating...

Recommended