Feb 12, 2026
Together AI launches production-grade orchestration for custom AI models with 1.4x–2.6x faster inference.
app The end of English-first voice AI, now on desktop Work, think, and use your voice in your own language.
February 12, 2026
app The end of English-first voice AI, now on desktop Work, think, and use your voice in your own language.
Official Decagon technical update and publication covering AI agents are never done: The new build-vs-buy calculus.
Tabular foundation models, such as TabPFNv2 and TabICL, have recently dethroned gradient-boosted trees at the top of predictive benchmarks, demonstrating the value of in-context learning for tabular data. We introduce TabICLv2, a new state-of-the-art foundation model for regression and classification built on three pillars: (1) a novel synthetic data generation engine designed for high pretraining
Tabular foundation models, such as TabPFNv2 and TabICL, have recently dethroned gradient-boosted trees at the top of predictive benchmarks, demonstrating the value of in-context learning for tabular data. We introduce TabICLv2, a new state-of-the-art foundation model for regression and classification built on three pillars: (1) a novel synthetic data generation engine designed for high pretraining
Official ElevenLabs technical update and publication covering Introducing ElevenLabs for Government.
Official ElevenLabs technical update and publication covering Klarna reduces Time to Resolution by 10X with ElevenAgents.
Algorithms & Theory
By Ryan Lopopolo, Member of the Technical Staff
Explore strategies for balancing quality and latency in real-time TTS AI models. Learn how Gradium achieves low-latency, high-quality speech synthesis for voice applications.
Official LMSYS Chatbot Arena release and benchmark update covering Unleashing Computational Power: Ultimate Latency Optimization of Qwen3 and Qwen3-VL on AMD MI300X Series.
Human-Computer Interaction and Visualization
Large language models (LLMs) often present answers with high apparent confidence despite lacking an explicit mechanism for reasoning about certainty or truth. While existing benchmarks primarily evaluate single-turn accuracy, truthfulness or confidence calibration, they do not capture how models behave when their responses are challenged in interactive settings. We introduce the Certainty Robustne
We built a feature that massively increased our internal token spend on Devin. But our PRs are now much more free of bugs and we can't go back.
Climate & Sustainability
Research papers point to the growing impact of Deep Think across fields
OpenAI for Government announces the deployment of a custom ChatGPT on GenAI.mil, bringing secure, safety-forward AI to U.S. defense teams.
Reproducibility is a cornerstone of science. FAIR (findable, accessible, interoperable, and reusable) data is often a vital step towards testing the reproducibility of results. The implementation of FAIR principles in the astrophysical simulation community is still varied. We approach the discussion of this topic mainly from a high-performance computing (HPC) point of view. We identify the main ob
Official technical announcement and publication from Hugging Face covering Transformers.js v4: Now Available on NPM!.
Get 6 months free access to Gradium's voice AI platform. 9M monthly credits, voice cloning, STT/TTS APIs for seed-funded startups building voice-first products.
OpenAI shares its approach to AI localization, showing how globally shared frontier models can be adapted to local languages, laws, and cultures without compromising safety.
Cartesia Is Hiring a SWE to Raise an Army of Claudes
What do language models generate when you don't tell them what to generate? New research reveals that LLM families have distinct 'knowledge priors'—GPT models default to code and math, Llama favors narratives, DeepSeek generates religious content, and Qwen outputs exam questions.
Sierra is heading into year three with over $150M in ARR, powered by rapid adoption from some of the world’s largest companies. The growth reflects a simple idea: when AI is built around real jobs to be done (not experiments), even the biggest enterprises can see meaningful impact fast.
After hearing feedback from Canadian founders in our network, we are adding Canada back to our list of accepted countries of incorporation.
Official technical announcement and publication from Hugging Face covering Introducing SyGra Studio.
Official Cursor (Anysphere) technical update and publication covering Read more →.
An autonomous lab combining OpenAI’s GPT-5 with Ginkgo Bioworks’ cloud automation cut cell-free protein synthesis costs by 40% through closed-loop experimentation.
OpenAI introduces Trusted Access for Cyber, a trust-based framework that expands access to frontier cyber capabilities while strengthening safeguards against misuse.
Education Innovation
OpenAI Frontier is an enterprise platform for building, deploying, and managing AI agents with shared context, onboarding, permissions, and governance.
GPT‑5.3-Codex is the most capable agentic coding model to date, combining the frontier coding performance of GPT‑5.2-Codex with the reasoning and professional knowledge capabilities of GPT‑5.2.
A speech recognition model purpose-built for low-latency voice interactions.
Acolad, the global leader in language and content solutions, and Gradium just announced a strategic partnership. The partnership reflects Acolad’s commitment to delivering secure, scalable, and governed AI-powered interpreting solutions, designed for enterprise and public-sector environments.
A speech recognition model purpose-built for low-latency voice interactions.
GPT-5.3-Codex is a Codex-native agent that pairs frontier coding performance with general reasoning to support long-horizon, real-world technical work.
Algorithms & Theory
Learn how to embed the Codex agent using the Codex App Server, a bidirectional JSON-RPC API powering streaming progress, tool use, approvals, and diffs.
Official ElevenLabs technical update and publication covering ElevenLabs raises $500M Series D at $11B valuation.
Official technical announcement and publication from Hugging Face covering Community Evals: Because we're done trusting black-box leaderboards over the community.
Official technical announcement and publication from Together AI covering Rime Arcana V3 Turbo and Rime Arcana V3 now available on Together AI.
Official Mistral AI technical update and publication covering Voxtral transcribes at the speed of sound..
Generative AI
Official technical announcement and publication from Hugging Face covering H Company's new Holo2 model takes the lead in UI Localization.
Official technical announcement and publication from Hugging Face covering The Future of the Global Open-Source AI Ecosystem: From DeepSeek to AI+.
Official technical announcement and publication from Hugging Face covering Training Design for Text-to-Image Models: Lessons from Ablations.
Hiring Alon Gavrielov further deepens Together AI’s commitment to building AI factories that deliver the most reliable, efficient, and scalable infrastructure for AI-native teams.
Discover the Sora feed philosophy—built to spark creativity, foster connections, and keep experiences safe with personalized recommendations, parental controls, and strong guardrails.
We propose RLAnything, a reinforcement learning framework that dynamically forges environment, policy, and reward models through closed-loop optimization, amplifying learning signals and strengthening the overall RL system for any LLM or agentic scenarios. Specifically, the policy is trained with integrated feedback from step-wise and outcome signals, while the reward model is jointly optimized vi
Official ElevenLabs technical update and publication covering Eleven v3 is Now Generally Available.
OpenAI and Snowflake partner in a $200M agreement to bring frontier intelligence into enterprise data, enabling AI agents and insights directly in Snowflake.
Feb 2, 2026
Fine-tuned open-source LLM judges can outperform GPT-5.2 at evaluating model outputs. Using Direct Preference Optimization on just 5,400 preference pairs, we trained GPT-OSS 120B to beat GPT-5.2 on human preference alignment—at 15x lower cost and 14x faster inference speeds.
Together Evaluations now supports OpenAI, Anthropic, and Google models for cross-provider benchmarking. Compare open-source, fine-tuned, and proprietary models side-by-side to make data-driven decisions on quality, cost, and performance—all in one platform.
Introducing the Codex app for macOS—a command center for AI coding and software development with multiple agents, parallel workflows, and long-running tasks.
OpenAI banned accounts linked to the Rybar network, some of which likely originated in Russia, that used AI to support multilingual influence activity across websites and social platforms.
OpenAI banned accounts that very likely originated in Cambodia and used AI to pose as recovery services, law firms, and authorities targeting people affected by fraud.
OpenAI banned accounts linked to a previously unreported, likely Russia-origin operation we dubbed "No Bell", using AI to produce criticism of the US and its allies for audiences across Africa.
OpenAI banned accounts using AI to support romance scam workflows, including outreach, translation, victim engagement, and investment-fraud lures.
OpenAI banned accounts that very likely originated in Cambodia and used AI for scam outreach to Indonesian loveseekers, including translation and engagement.
OpenAI banned an account linked to an individual associated with Chinese law enforcement, using AI to plan influence activity, harassment, and online operations.
OpenAI banned accounts linked to a previously unreported operation we dubbed "Trolling Stone", using AI to generate comments about an alleged Russian cult leader’s arrest in Argentina.
OpenAI banned likely China-origin accounts using AI to research US persons, locations, and social-engineering tactics.
Nash regret has recently emerged as a principled fairness-aware performance metric for stochastic multi-armed bandits, motivated by the Nash Social Welfare objective. Although this notion has been extended to linear bandits, existing results suffer from suboptimality in ambient dimension $d$, stemming from proof techniques that rely on restrictive concentration inequalities. In this work, we resol
Professor Sara Bernardini has been named as one of the winners of the Suffrage Science Awards in Maths and Computing.
Google AI Ultra subscribers in the U.S. can try out Project Genie, an experimental research prototype that lets you create and explore worlds.
Serving Large Language Models (LLMs) under mixed workloads--short, latency-sensitive interactive queries alongside long, throughput-oriented batch requests--poses a fundamental scheduling challenge. Standard First-Come, First-Served (FCFS) policies suffer from severe head-of-line blocking, leading to high tail latency and underutilized hardware. We introduce EWSJF (Effective Workload-based Shortes
Official ElevenLabs technical update and publication covering We are on the grid.
Professor Leslie Ann Goldberg, Head of the Department of Computer Science, has been appointed to the inaugural cohort of Fellows of the Academy for the Mathematical Sciences. She joins a distinguished group of UK-based mathematicians working across academia, education, business, industry and government.
How OpenAI built an in-house AI data agent that uses GPT-5, Codex, and memory to reason over massive datasets and deliver reliable insights in minutes.
On February 13, 2026, alongside the previously announced retirement of GPT‑5 (Instant, Thinking, and Pro), we will retire GPT‑4o, GPT‑4.1, GPT‑4.1 mini, and OpenAI o4-mini from ChatGPT. In the API, there are no changes at this time.
Official technical announcement and publication from Hugging Face covering Introducing Daggr: Chain apps programmatically, inspect visually.
We’re launching Soniox v4 Async, a major leap forward in speech recognition.
We’re launching Soniox v4 Async, a major leap forward in speech recognition.
Taisei Corporation’s HR team is leading the rollout of ChatGPT Enterprise to drive AI-powered talent development across the organization.
Official ElevenLabs technical update and publication covering Revolut selects ElevenLabs Agents to bolster customer support.
Generative AI
OpenAI launches the EU Economic Blueprint 2.0 with new data, partnerships, and initiatives to accelerate AI adoption, skills, and growth across Europe.
Apply for the EMEA Youth & Wellbeing Grant, a €500,000 program funding NGOs and researchers advancing youth safety and wellbeing in the age of AI.
Learn how OpenAI protects user data when AI agents open links, preventing URL-based data exfiltration and prompt injection with built-in safeguards.
Cognition expands to Europe and opens a London office.
Cognizant has partnered with Cognition to deploy Devin and Windsurf across its engineering teams and customer base.
Official technical announcement and publication from Hugging Face covering We Got Claude to Build CUDA Kernels and teach open models!.
Gradium's voice AI technology powers Invincible Voice, an open-source assistive system helping people with ALS and speech loss communicate in real-time.
January 28, 2026
January 28, 2026
Generative AI
Official technical announcement and publication from Hugging Face covering Architectural Choices in China's Open-Source AI Ecosystem: Building Beyond DeepSeek.
Official technical announcement and publication from Hugging Face covering Alyah ⭐️: Toward Robust Evaluation of Emirati Dialect Capabilities in Arabic LLMs.
PVH Corp., parent company of Calvin Klein and Tommy Hilfiger, is adopting ChatGPT Enterprise to bring AI into fashion design, supply chain, and consumer engagement.
Official technical announcement and publication from Hugging Face covering Unlocking Agentic RL Training for GPT-OSS: A Practical Retrospective.
Prism is a free LaTeX-native workspace with GPT-5.2 built in, helping researchers write, collaborate, and reason in one place.
TRUSTBANK partnered with Recursive to build Choice AI using OpenAI models, enabling personalized conversational recommendations that simplify Furusato Nozei gift discovery.
Official Mistral AI technical update and publication covering Terminally online Mistral Vibe..
Posted on January 27, 2026
Introducing DSGym—a holisti evaluation and training framework for LLM-based data science agents. Features 90+ bioinformatics tasks, 92 Kaggle competitions, and synthetic trajectory generation. Our 4B model achieves state-of-the-art performance among open-source models through exe
Indeed’s CRO Maggie Hulce shares how AI is transforming job search, recruiting, and talent acquisition for employers and job seekers.
6584 articles sourced historically · 100 per page