Dolphin 3.0 Released (Llama 3.1 + 3.2 + Qwen 2.5): A Local-First, Steerable AI...
Artificial intelligence has come a long way, transforming the way we work,...
Graph Generative Pre-trained Transformer (G2PT): An Auto-Regressive Model Designed to Learn Graph Structures through...
Graph generation is an important task across various fields, including molecular design...
From Latent Spaces to State-of-the-Art: The Journey of LightningDiT
Latent diffusion models are advanced techniques for generating high-resolution images by compressing...
Enhancing Protein Docking with AlphaRED: A Balanced Approach to Protein Complex Prediction
Protein docking, the process of predicting the structure of protein-protein complexes, remains...
Scaling of Search and Learning: A Roadmap to Reproduce o1 from Reinforcement Learning Perspective
Achieving expert-level performance in complex reasoning tasks is a significant challenge in...
Researchers from NVIDIA, CMU and the University of Washington Released ‘FlashInfer’: A Kernel Library...
Large Language Models (LLMs) have become an integral part of modern AI...
PRIME: An Open-Source Solution for Online Reinforcement Learning with Process Rewards to Advance Reasoning...
Large Language Models (LLMs) face significant scalability limitations in improving their reasoning...
FutureHouse Researchers Propose Aviary: An Extensible Open-Source Gymnasium for Language Agents
Artificial intelligence (AI) has made significant strides in developing language models capable...
This AI Paper Introduces SWE-Gym: A Comprehensive Training Environment for Real-World Software Engineering Agents
Software engineering agents have become essential for managing complex coding tasks, particularly...
Meta AI Introduces EWE (Explicit Working Memory): A Novel Approach that Enhances Factuality in...
Large Language Models (LLMs) have revolutionized text generation capabilities, but they face...























