Official technical announcement and publication from Hugging Face covering LeRobot v0.5.0: Scaling Every Dimension.
Official technical announcement and publication from Hugging Face covering Ulysses Sequence Parallelism: Training with Million-Token Contexts.
Posted on March 9, 2026
We present AutoResearch-RL, a framework in which a reinforcement learning agent conducts open-ended neural architecture and hyperparameter research without human supervision, running perpetually until a termination oracle signals convergence or resource exhaustion. At each step the agent proposes a code modification to a target training script, executes it under a fixed wall clock time budget, obs
Recent generative video world models aim to simulate visual environment evolution, allowing an observer to interactively explore the scene via camera control. However, they implicitly assume that the world only evolves within the observer's field of view. Once an object leaves the observer's view, its state is "frozen" in memory, and revisiting the same region later often fails to reflect events t
Natural Language Processing
Climate & Sustainability
Codex Security is an AI application security agent that analyzes project context to detect, validate, and patch complex vulnerabilities with higher confidence and less noise.
By combining rigorous model evaluation, full-platform use of OpenAI, and agent workflows, Balyasny is reinventing investment research.
Using OpenAI reasoning models, Descript unlocked automatic localization of large content libraries without losing timing or meaning.
Estimating heterogeneous treatment effects (HTEs) from right-censored survival data is critical in high-stakes applications such as precision medicine and individualized policy-making. Yet, the survival analysis setting poses unique challenges for HTE estimation due to censoring, unobserved counterfactuals, and complex identification assumptions. Despite recent advances, from Causal Survival Fores
Official technical announcement and publication from Hugging Face covering Bringing Robotics AI to Embedded Platforms: Dataset Recording, VLA Fine‑Tuning, and On‑Device Optimizations.
Introducing GPT-5.4, OpenAI’s most most capable and efficient frontier model for professional work, with state-of-the-art coding, computer use, tool search, and 1M-token context.
Official technical announcement and publication from OpenAI covering GPT-5.4 Thinking System Card.
OpenAI introduces CoT-Control and finds reasoning models struggle to control their chains of thought, reinforcing monitorability as an AI safety safeguard.
OpenAI shares new tools, certifications, and measurement resources to help schools and universities close AI capability gaps and expand opportunity.
OpenAI introduces ChatGPT for Excel and new financial app integrations, powered by GPT-5.4 to accelerate modeling, research, and analysis in regulated environments.
Official technical announcement and publication from Hugging Face covering Introducing Modular Diffusers - Composable Building Blocks for Diffusion Pipelines.
Practical insights and frameworks to turn AI progress into business advantage
As GPU throughput outpaces memory bandwidth, kernels must evolve. We introduce FlashAttention-4, featuring new pipelining for maximum overlap, 2-CTA MMA modes to reduce shared memory traffic, and a hardware-software hybrid approach to softmax exponentials.
By focusing on people, not pilots, the Bundesliga club is scaling efficiency, creativity, and knowledge—without losing its football identity.
Five AI value models show how leaders can sequence AI from workforce fluency to process reinvention and build durable business advantage.
At AI Native Conf, Together AI announced breakthroughs across kernels, RL, and inference optimization — including FlashAttention-4, ThunderAgent, and together.compile. Research that ships to production. That's the AI Native Cloud.
Generative AI
A new preprint extends single-minus amplitudes to gravitons, with GPT-5.2 Pro helping derive and verify nonzero graviton tree amplitudes in quantum gravity.
Serving long prompts doesn't have to mean slow responses. Learn how Together AI's CPD architecture separates warm and cold inference workloads to deliver 40% higher throughput and dramatically lower time-to-first-token for long-context LLM serving.
Axios COO Allison Murphy explains how the company uses AI to support local reporters, streamline newsroom workflows, and deliver high-impact local journalism at scale.
OpenAI introduces the Learning Outcomes Measurement Suite to assess AI’s impact on student learning across diverse educational environments over time.
Mar 4, 2026
March 04, 2026
Posted on March 4, 2026
Official technical announcement and publication from Hugging Face covering PRX Part 3 — Training a Text-to-Image Model in 24h!.
Gemini 3.1 Flash-Lite is our fastest and most cost-efficient Gemini 3 series model yet.
DPhil student Zandi Eberstadt on expanding our conceptual framework for detecting non-human intelligence.
Official technical announcement and publication from OpenAI covering GPT-5.3 Instant System Card.
Official technical announcement and publication from OpenAI covering GPT-5.3 Instant: Smoother, more useful everyday conversations.
Mar 3, 2026
We've refreshed our visual identity — designed with Pentagram to express how Together AI connects open-source innovation, systems research, and builders to unlock new possibilities.
Cursor Recurring Revenue Doubles in Three Months to $2 Billion
Posted on March 2, 2026
We are sharing an early preview of our ongoing SWE-1.6 training run.
Details on OpenAI’s contract with the Department of War, outlining safety red lines, legal protections, and how AI systems will be deployed in classified environments.
Stateful Runtime for Agents in Amazon Bedrock brings persistent orchestration, memory, and secure execution to multi-step AI workflows powered by OpenAI.
Today we’re announcing $110B in new investment at a $730B pre money valuation. This includes $30B from SoftBank, $30B from NVIDIA, and $50B from Amazon.
OpenAI and Amazon announce a strategic partnership bringing OpenAI’s Frontier platform to AWS, expanding AI infrastructure, custom models, and enterprise AI agents.
Microsoft and OpenAI continue to work closely across research, engineering, and product development, building on years of deep collaboration and shared success.
OpenAI shares updates on its mental health safety work, including parental controls, trusted contacts, improved distress detection, and recent litigation developments.
Devin is a cloud agent platform for engineering teams. You work with it like a teammate — give it tasks, review its PRs, and let it handle your backlog. Here's how we use it to build Devin itself.
Our latest image generation model offers advanced world knowledge, production ready specs, subject consistency and more, all at Flash speed.
The University of Oxford has launched Oxford × QRT Labs as part of a major new long-term philanthropic partnership with Qube Research & Technologies (QRT), alongside Imperial College London and the University of Cambridge.
OpenAI and Pacific Northwest National Laboratory introduce DraftNEPABench, a new benchmark evaluating how AI coding agents can accelerate federal permitting—showing potential to reduce NEPA drafting time by up to 15% and modernize infrastructure reviews.
OpenAI and Figma launch a new Codex integration that connects code and design, enabling teams to move between implementation and the Figma canvas to iterate and ship faster.
Official technical announcement and publication from Hugging Face covering Mixture of Experts (MoEs) in Transformers.
Vector Institute’s third annual Remarkable 2026 conference brought together over 1,500 researchers and industry leaders in person and online on February 19-20 to explore how AI research translates into real-world […] The post Remarkable 2026 Poster Session: 60 research projects shaping AI’s future appeared first on Vector Institute for Artificial Intelligence .
Today, we launch Cognition for Government to modernize America’s critical infrastructure with AI software engineering.
Official technical announcement and publication from Together AI covering CoderForge-Preview: SOTA open dataset for training efficient coding agents.
Our latest threat report examines how malicious actors combine AI models with websites and social platforms—and what it means for detection and defense.
We are open-sourcing the initial version of RCCLX – an enhanced version of RCCL that we developed and tested on Meta’s internal workloads. RCCLX is fully integrated with Torchcomms and aims to empower researchers and developers to accelerate innovation, regardless of their chosen backend. Communication patterns for AI models are constantly evolving, as are hardware [...] Read More... The post RCCL
Vector researchers developed CRISPNAM-FG, a trustworthy AI model that predicts the risk of developing diabetes-related foot complications for patients discharged from hospitals while providing complete transparency in how each decision […] The post CRISPNAM-FG: An interpretable Fine-Gray deep survival model for competing risks in health care appeared first on Vector Institute for Artificial Intell
OpenAI appoints Arvind KC as Chief People Officer to help scale the company, strengthen its culture, and lead how work evolves in the age of AI.
Today we’re releasing Devin 2.2, the most important update to Devin since launch.
By Daniel Kitts “Where were those 10 years ago in AI?” said Stephen Southin to the crowd at Vector’s first Demo Day. The AI industry veteran was energized after hearing […] The post Demo Day: How the Vector Institute helps Canadian startups turn innovative ideas into commercial reality appeared first on Vector Institute for Artificial Intelligence .
SWE-bench Verified is increasingly contaminated and mismeasures frontier coding progress. Our analysis shows flawed tests and training leakage. We recommend SWE-bench Pro.
OpenAI announces Frontier Alliance Partners to help enterprises move from AI pilots to production with secure, scalable agent deployments.
State-of-the-art speech models like Whisper and Deepgram score near-human on benchmarks — then fail 39% of the time on street names. New research from Together AI exposes the gap and a fix.
February 23, 2026
Managing the data and metadata during the active development phase of an experimental project presents a significant challenge, particularly in collaborative research. This phase is frequently overlooked in Data Management Plans included in project proposals, despite its important role in ensuring reproducibility and preventing the need for retroactive reconstruction at the time of publication. He
We share our AI model’s proof attempts for the First Proof math challenge, testing research-grade reasoning on expert-level problems.
Official technical announcement and publication from Hugging Face covering Train AI models with Unsloth and Hugging Face Jobs for FREE.
Official technical announcement and publication from Hugging Face covering GGML and llama.cpp join HF to ensure the long-term progress of Local AI.
Official LMSYS Chatbot Arena release and benchmark update covering Unlocking 25x Inference Performance with SGLang on NVIDIA GB300 NVL72.
3.1 Pro is designed for tasks where a simple answer isn’t enough.
OpenAI commits $7.5M to The Alignment Project to fund independent AI alignment research, strengthening global efforts to address AGI safety and security risks.
Introducing the new Soniox SDKs for Python, Node, Web, React, and React Native.
Standard diffusion language models can't use KV caching and need too many refinement steps to be practical. CDLM fixes both with a post-training recipe that enables exact block-wise KV caching and trajectory-consistent step reduction — delivering up to 14.5x latency improvements
Law firms and professional service networks have been using Harvey to build new service models and add value collaboratively.
Official LMSYS Chatbot Arena release and benchmark update covering Deploying DeepSeek on GB300 NVL72: Big Wins in Long-Context Inference.
February 19, 2026
Introducing the new Soniox SDKs for Python, Node, Web, React, and React Native.
OpenAI for India expands AI access across the country—building local infrastructure, powering enterprises, and advancing workforce skills.
Official technical announcement and publication from Hugging Face covering IBM and UC Berkeley Diagnose Why Enterprise Agents Fail Using IT-Bench and MAST.
The Gemini app now features our most advanced music generation model Lyria 3, empowering anyone to make 30-second tracks using text or images.
OpenAI and Paradigm introduce EVMbench, a benchmark evaluating AI agents’ ability to detect, patch, and exploit high-severity smart contract vulnerabilities.
Official technical announcement and publication from Hugging Face covering One-Shot Any Web App with Gradio's gr.HTML.
Official Index Ventures technical update and publication covering From Assistive AI to Authoritative Systems: How Wonderful is taking Enterprise AI out of Pilot Mode.
Machine Perception
Google DeepMind brings National Partnerships for AI initiative to India, scaling AI for science and education
InteractionLabs, the company behind the Ongo living lamp robot, and Gradium announce a partnership to bring expressive, real-time voice AI to robotics.
Official Groq (LPUs) technical update and publication covering GroqCloud: Expanding to Meet Demand.
In a comprehensive study by Daily (Pipecat), Soniox was recognized as a top-tier provider for real-time voice agents.
Official LMSYS Chatbot Arena release and benchmark update covering SGLang-Diffusion: Advanced Optimizations for Production-Ready Video Generation.
In a comprehensive study by Daily (Pipecat), Soniox was recognized as a top-tier provider for real-time voice agents.
A new preprint shows GPT-5.2 proposing a new formula for a gluon amplitude, later formally proved and verified by OpenAI and academic collaborators.
Introducing Lockdown Mode and Elevated Risk labels in ChatGPT to help organizations defend against prompt injection and AI-driven data exfiltration.
GABRIEL is a new open-source toolkit from OpenAI that uses GPT to turn qualitative text and images into quantitative data, helping social scientists analyze research at scale.
How OpenAI built a real-time access system combining rate limits, usage tracking, and credits to power continuous access to Sora and Codex.
Official technical announcement and publication from Hugging Face covering Custom Kernels for All from Codex and Claude.
Our most specialized reasoning mode is now updated to solve modern science, research and engineering challenges.
Introducing GPT-5.3-Codex-Spark—our first real-time coding model. 15x faster generation, 128k context, now in research preview for ChatGPT Pro users.
Together AI launches production-grade orchestration for custom AI models with 1.4x–2.6x faster inference.
6584 articles sourced historically · 100 per page