Hey there, friend using a screen reader. I've carefully labeled every button and image on this site. Hope you have a smooth experience. If anything is inconvenient, my email is in the footer — just let me know and I'll fix it. — Store Owner
Every item keeps its source, verification time, and eligibility notes. Filter by direction, difficulty, format, and deadline to find something you can act on now.
An environment gives an agent a task, responds to its actions with observations, and scores the outcome. The resulting rewards can measure an agent's performance during evaluation or provide a learning signal during training. For an introduction to this interaction loop, see our blogpost on environments . Within the e…
Hugging FacePublished Sep 28, 2026
Cross-checked · Hugging Face BlogVerified Oct 9, 2026
Today, we’re introducing two new models that bring advancements in near real-time reasoning to more effectively enable voice agents and make conversing with AI feel more intuitive and intelligent.
Google DeepMindPublished Sep 16, 2026
Official source · Google DeepMind BlogVerified Oct 9, 2026
A rewrite this size wasn't affordable before agents. Here's what porting the Copilot agent runtime to 800,000 lines of production Rust actually took. The post Migrating the GitHub Copilot runtime to Rust, using Copilot appeared first on The GitHub Blog .
GitHubPublished Sep 17, 2026
Official source · GitHub AI & MLVerified Oct 9, 2026
Since launching a year ago, the Microsoft Research Asia — Singapore lab has established a strong foundation, deepened collaboration across government, academia, and industry, and explored how frontier AI research can create real-world value. The post One year in: How Microsoft Research Asia – Singapore is advancing re…
Microsoft ResearchPublished Sep 29, 2026
Official source · Microsoft ResearchVerified Oct 9, 2026
Our latest paper, ProvenanceGuard: Source-Aware Factuality Verification for MCP-Based LLM Agents (read it on Hugging Face , or on arXiv in the meantime), targets that gap. The failure mode we care about is one we call cross-source conflation: a claim that is true somewhere in the evidence, but attributed to the wrong…
Hugging FacePublished Sep 29, 2026
Cross-checked · Hugging Face BlogVerified Oct 9, 2026
Today, we’re announcing our new frontier model, Gemini 4 Argon, which is rolling out to a set of trusted cyber defenders through our Fairwind Program . Built to sustain deep reasoning across complex, long-horizon workflows, Argon is fundamentally changing the way we work and build at Google. It delivers frontier perfo…
Google DeepMindPublished Oct 1, 2026
Official source · Google DeepMind BlogVerified Oct 9, 2026
Learn to direct AI agents, critically review their output, and keep technical judgment at the center of your workflow. The post AI is changing developer work. Here are three skills to strengthen. appeared first on The GitHub Blog .
GitHubPublished Oct 2, 2026
Official source · GitHub AI & MLVerified Oct 9, 2026
Skala 1.1, the updated deep-learning exchange-correlation functional from Microsoft Research, provides greater accuracy, expanded accessibility across the computational chemistry ecosystem, and a living benchmark to track computational performance. The post Broadening access to Skala creates a faster path to predictiv…
Microsoft ResearchPublished Aug 21, 2026
Official source · Microsoft ResearchVerified Oct 9, 2026
Evaluation, however, hasn't kept pace: it remains fragmented and unstandardized. The gold standard is human preference scores such as MOS or MUSHRA (more on metrics ). To this end, several arena-based leaderboards have established themselves as useful reference points for the community:
Hugging FacePublished Sep 30, 2026
Cross-checked · Hugging Face BlogVerified Oct 9, 2026
Building on the momentum of last week's Gemini 3.8 Live launch, today we are excited to introduce Gemini 3.8 Live with Live Avatar — bringing near real-time visual presence to our native live dialogue models. By pairing near real-time video generation with speech, the Live Avatar feature creates an experience that lis…
Google DeepMindPublished Sep 25, 2026
Official source · Google DeepMind BlogVerified Oct 9, 2026
What is a developer to do when they need something more tangible than a chat box? Enter canvases. The post When chat is the wrong UI appeared first on The GitHub Blog .
GitHubPublished Sep 25, 2026
Official source · GitHub AI & MLVerified Oct 9, 2026
Training AI agents with reinforcement learning can be challenging because their tools, context, and decision-making are managed by complex frameworks. Agent Lightning connects existing agents to RL training, making it easier to improve them without rebuilding them. The post Agent Lightning v1.0: A 3,500-Line Lightweig…
Microsoft ResearchPublished Oct 8, 2026
Official source · Microsoft ResearchVerified Oct 9, 2026
We dive into these questions and other AI hot takes on the latest episode of the GitHub Podcast. The post Should you read the code, is RAG dead, and did Skills kill MCP? appeared first on The GitHub Blog .
GitHubPublished Sep 18, 2026
Official source · GitHub AI & MLVerified Oct 9, 2026