AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,959 results
24 Apr 2026

Introducing Vision-Language Reinforcement Learning in SkyRL

ToolsDGX agent

SkyRL is a reinforcement learning framework from Anyscale that integrates vision-language models to enable agents to learn from visual observations and natural language instructions. The approach comb

New Frontier LLM Orchestrator Model @SakanaAILabs The fugu model series provides a novel paradigm for test-time scaling. The models are trai…

AgentsDGX agent

New Frontier LLM Orchestrator Model @SakanaAILabs The fugu model series provides a novel paradigm for test-time scaling. The models are trained to call LLMs and infer not only which individual model i

pi + @ollama + gemma4 + @p0 - each very useful individually, but so much better when composed together!

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

pi + @ollama + gemma4 + @p0 - each very useful individually, but so much better when composed together! We built a completely free CLI agent with @badlogicgames's Pi agent, @ollama (Gemma 4), and Para

Planetary Exploration 3.0: A Roadmap for Software-Defined, Radically Adaptive Space Systems

AgentsDGX agent

arXiv:2604.20910v1 Announce Type: cross Abstract: The surface and subsurface of worlds beyond Mars remain largely unexplored. Yet these worlds hold keys to fundamental questions in planetary science -

Stratified Topological Autonomy for Long-Range Coordination (STALC)

AgentsDGX agent

arXiv:2503.10475v4 Announce Type: replace Abstract: In this paper, we present Stratified Topological Autonomy for Long-Range Coordination (STALC), a hierarchical planning approach for multi-robot coor

We’ve been using Sakana Fugu internally for our own research and coding. Instead of relying on a single model, it dynamically orchestrates t…

AgentsDGX agent

We’ve been using Sakana Fugu internally for our own research and coding. Instead of relying on a single model, it dynamically orchestrates the best combination of open and closed models for any task.

23 Apr 2026

CubeDAgger: Interactive Imitation Learning for Dynamic Systems with Efficient yet Low-risk Interaction

SafetyDGX agent

arXiv:2505.04897v2 Announce Type: replace-cross Abstract: Interactive imitation learning makes an agent's control policy robust by stepwise supervisions from an expert. The recent algorithms mostly em

Great food, even better company!

AgentsDGX agent

Great food, even better company! Last night, we teamed up with First Commits to co-host our first dinner together for founding engineers at Penny Roma. We shared honest conversations and real stories

LEAD: Breaking the No-Recovery Bottleneck in Long-Horizon Reasoning

Local AiDGX agent

arXiv:2603.06870v2 Announce Type: replace Abstract: Long-horizon execution in Large Language Models (LLMs) remains unstable even when high-level strategies are provided. Evaluating on controlled algor

Lifecycle-Aware Federated Continual Learning in Mobile Autonomous Systems

AgentsDGX agent

arXiv:2604.20745v1 Announce Type: cross Abstract: Federated continual learning (FCL) allows distributed autonomous fleets to adapt collaboratively to evolving terrain types across extended mission lif

NanoCockpit: Performance-optimized Application Framework for AI-based Autonomous Nanorobotics

AgentsDGX agent

arXiv:2601.07476v2 Announce Type: replace Abstract: Autonomous nano-drones, powered by vision-based tiny machine learning (TinyML) models, are a novel technology gaining momentum thanks to their broad

Round 2 for Max Agency! The Listen team know what they're doing. Check it out 👇

AgentsDGX agent

Round 2 for Max Agency! The Listen team know what they're doing. Check it out 👇 🎙️ Talked to @ListenLabs co-founder + CTO @florian_jue in the latest Max Agency. Really enjoyed hearing about the archit

X-Cache: Cross-Chunk Block Caching for Few-Step Autoregressive World Models Inference

AgentsDGX agent

arXiv:2604.20289v1 Announce Type: new Abstract: Real-time world simulation is becoming a key infrastructure for scalable evaluation and online reinforcement learning of autonomous driving systems. Rec

22 Apr 2026

CityRAG: Stepping Into a City via Spatially-Grounded Video Generation

AgentsDGX agent

arXiv:2604.19741v1 Announce Type: new Abstract: We address the problem of generating a 3D-consistent, navigable environment that is spatially grounded: a simulation of a real location. Existing video

InsideOut: Measuring and Mitigating Insider-Outsider Bias in Interview Script Generation

Model ReleasesDGX agent

arXiv:2509.21080v2 Announce Type: replace-cross Abstract: Advancements in Large language models (LLMs) have enabled a variety of downstream applications like story and interview script generation. How

Learn more about this model release https://x.com/Alibaba_Qwen/status/2046939764428009914?s=20

AgentsDGX agent

Learn more about this model release https://x.com/Alibaba_Qwen/status/2046939764428009914?s=20 🚀 Meet Qwen3.6-27B, our latest dense, open-source model, packing flagship-level coding power! Yes, 27B, a

OpenAI is shutting down text-embedding-3-small?!? I strongly believe that if you shut down a closed-source embedding model that you should o…

AgentsDGX agent

OpenAI is shutting down text-embedding-3-small?!? I strongly believe that if you shut down a closed-source embedding model that you should open-source. Imaging the trillions of tokens that will no lon

Sony AI says its autonomous ping pong robot is the first robot to attain expert-level performance in a physical sport after beating some top-level human players (Will Dunham/Reuters)

AgentsDGX agent

Will Dunham / Reuters: Sony AI says its autonomous ping pong robot is the first robot to attain expert-level performance in a physical sport after beating some top-level human players — An autonomous

The Triadic Loop: A Framework for Negotiating Alignment in AI Co-hosted Livestreaming

SafetyDGX agent

arXiv:2604.18850v1 Announce Type: cross Abstract: AI systems are increasingly embedded in multi-user social environments, yet most alignment frameworks conceptualize interaction as a dyadic relationsh

Time Series Augmented Generation for Financial Applications

Model ReleasesDGX agent

arXiv:2604.19633v1 Announce Type: new Abstract: Evaluating the reasoning capabilities of Large Language Models (LLMs) for complex, quantitative financial tasks is a critical and unsolved challenge. St

Two new TPUs to power the next wave of AI training and inference at Google

HardwareDGX agent

Google LLC introduced two new custom silicon chips for artificial intelligence today at Google Cloud Next 2026, unveiling two distinct Tensor Processor Unit architectures built for training and infere

What Makes an LLM a Good Optimizer? A Trajectory Analysis of LLM-Guided Evolutionary Search

Local AiDGX agent

arXiv:2604.19440v1 Announce Type: new Abstract: Recent work has demonstrated the promise of orchestrating large language models (LLMs) within evolutionary and agentic optimization systems. However, th

21 Apr 2026

A Rapid Deployment Pipeline for Autonomous Humanoid Grasping Based on Foundation Models

AgentsDGX agent

arXiv:2604.17258v1 Announce Type: new Abstract: Deploying a humanoid robot to manipulate a new object has traditionally required one to two days of effort: data collection, manual annotation, 3D model

Automatic Slide Updating with User-Defined Dynamic Templates and Natural Language Instructions

Model ReleasesDGX agent

arXiv:2604.17894v1 Announce Type: new Abstract: Presentation slides are a primary medium for data-driven reporting, yet keeping complex, analytics-style decks up to date remains labor-intensive. Exist

From keynote to the terminal: Join our Next ‘26 developer livestreams

Model ReleasesDGX agent

The main stage at Google Cloud Next is where the vision is set. This year, we’re bridging the gap between those massive 'Cloud-scale' announcements and your local terminal. We are thrilled to announce

memory ended not being valuable in ChatGPT possible this exact implementation isn't valuable either but some memory implementation will be v…

AgentsDGX agent

memory ended not being valuable in ChatGPT possible this exact implementation isn't valuable either but some memory implementation will be valuable, will create a bunch of lock in. and everyone is rus

need open standards

AgentsDGX agent

The post likely discusses the importance of open standards in AI and software development, advocating for transparency and interoperability rather than proprietary solutions. Harrison Chase, co-founde

OneDrive: Unified Multi-Paradigm Driving with Vision-Language-Action Models

AgentsDGX agent

arXiv:2604.17915v1 Announce Type: new Abstract: Vision-Language Models(VLMs) excel at autoregressive text generation, yet end-to-end autonomous driving requires multi-task learning with structured out

probably true (and agree on it being open)

AgentsDGX agent

This post by Harrison Chase (creator of LangChain) likely discusses the concept of statements or propositions that are 'probably true' and advocates for keeping such discussions or determinations open

Systematic Capability Benchmarking of Frontier Large Language Models for Offensive Cyber Tasks

Model ReleasesDGX agent

arXiv:2604.17159v1 Announce Type: cross Abstract: We present, to our knowledge, the most comprehensive cross-model evaluation of LLM agents on offensive cybersecurity tasks, benchmarking 10 frontier m

Think before Go: Hierarchical Reasoning for Image-goal Navigation

SafetyDGX agent

arXiv:2604.17407v1 Announce Type: new Abstract: Image-goal navigation steers an agent to a target location specified by an image in unseen environments. Existing methods primarily handle this task by

XEmbodied: A Foundation Model with Enhanced Geometric and Physical Cues for Large-Scale Embodied Environments

AgentsDGX agent

arXiv:2604.18484v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models drive next-generation autonomous systems, but training them requires scalable, high-quality annotations from complex

20 Apr 2026

Facial-Expression-Aware Prompting for Empathetic LLM Tutoring

Model ReleasesDGX agent

arXiv:2604.15336v1 Announce Type: cross Abstract: Large language models (LLMs) enable increasingly capable tutoring-style conversational agents, yet effective tutoring requires sensitivity to learners

Fuzzy Logic Theory-based Adaptive Reward Shaping for Robust Reinforcement Learning (FARS)

AgentsDGX agent

arXiv:2604.15772v1 Announce Type: new Abstract: Reinforcement learning (RL) often struggles in real-world tasks with high-dimensional state spaces and long horizons, where sparse or fixed rewards seve

good point about another benefit of open memory

AgentsDGX agent

good point about another benefit of open memory @hwchase17 Memory being open isn't just about lock in, it's about auditability. If a bad memory gets injected, you need to be able to see it, revoke it,

Kimi has been the most popular model on Fireworks, both out of the box and as a fine-tuning base (including Composer 2) Now Kimi K2.6 is liv…

ToolsDGX agent

Kimi has been the most popular model on Fireworks, both out of the box and as a fine-tuning base (including Composer 2) Now Kimi K2.6 is live with huge jumps (10+%) in coding, long-running agents and

kimi k2.6 is now 'available' on ollama cloud

Local AiDGX agent

Kimi K2.6 is an open-source model featuring advanced coding, long-horizon execution, and agent swarm capabilities that is now available via Ollama Cloud . The model excels in coding and agentic tools

{pi}_{0.7}: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

AgentsDGX agent

arXiv:2604.15483v1 Announce Type: new Abstract: We present a new robotic foundation model, called {pi}_{0.7}, that can enable strong out-of-the-box performance in a wide range of scenarios. {pi}_{0.7}

The Jensen + @dwarkesh_sp podcast was fantastic. Jensen is someone who understood how ecosystems work and someone who understands real-world…

Model ReleasesDGX agent

The Jensen + @dwarkesh_sp podcast was fantastic. Jensen is someone who understood how ecosystems work and someone who understands real-world trade, policy and controls work. And in some deeper sense h

Zero-Shot Scalable Resilience in UAV Swarms: A Decentralized Imitation Learning Framework with Physics-Informed Graph Interactions

SafetyDGX agent

arXiv:2604.15762v1 Announce Type: new Abstract: Large-scale Unmanned Aerial Vehicle (UAV) failures can split an unmanned aerial vehicle swarm network into disconnected sub-networks, making decentraliz

19 Apr 2026

LangChain Community Spotlight: Saving $1M in LLM Costs 💰 Gustaf, an AI Engineer, shows how he reduced a production RAG chatbot's costs by 9…

ApplicationsDGX agent

LangChain Community Spotlight: Saving $1M in LLM Costs 💰 Gustaf, an AI Engineer, shows how he reduced a production RAG chatbot's costs by 90% and improved latency by 82% through caching, intelligent r

17 Apr 2026

ChatSVA: Bridging SVA Generation for Hardware Verification via Task-Specific LLMs

Model ReleasesDGX agent

arXiv:2604.02811v2 Announce Type: replace-cross Abstract: Functional verification consumes over 50% of the IC development lifecycle, where SystemVerilog Assertions (SVAs) are indispensable for formal

Conformal Policy Control

SafetyDGX agent

arXiv:2603.02196v2 Announce Type: replace-cross Abstract: An agent must try new behaviors to explore and improve. In high-stakes environments, an agent that violates safety constraints may cause harm

Dual Pose-Graph Semantic Localization for Vision-Based Autonomous Drone Racing

AgentsDGX agent

arXiv:2604.15168v1 Announce Type: new Abstract: Autonomous drone racing demands robust real-time localization under extreme conditions: high-speed flight, aggressive maneuvers, and payload-constrained

EviSearch: A Human in the Loop System for Extracting and Auditing Clinical Evidence for Systematic Reviews

Model ReleasesDGX agent

arXiv:2604.14165v1 Announce Type: new Abstract: We present EviSearch, a multi-agent extraction system that automates the creation of ontology-aligned clinical evidence tables directly from native tria

One-shot learning for the complex dynamical behaviors of weakly nonlinear forced oscillators

AgentsDGX agent

arXiv:2604.15181v1 Announce Type: new Abstract: Extrapolative prediction of complex nonlinear dynamics remains a central challenge in engineering. This study proposes a one-shot learning method to ide

16 Apr 2026

Bi-Predictability: A Real-Time Signal for Monitoring LLM Interaction Integrity

AgentsDGX agent

arXiv:2604.13061v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed in high-stakes autonomous and interactive workflows, where reliability demands continuous, multi-

Co-FactChecker: A Framework for Human-AI Collaborative Claim Verification Using Large Reasoning Models

AgentsDGX agent

arXiv:2604.13706v1 Announce Type: new Abstract: Professional fact-checkers rely on domain knowledge and deep contextual understanding to verify claims. Large language models (LLMs) and large reasoning

GPT-Rosalind, our Life Sciences model series, is optimized for scientific workflows, with stronger performance in protein and chemical reaso…

AgentsDGX agent

GPT-Rosalind, our Life Sciences model series, is optimized for scientific workflows, with stronger performance in protein and chemical reasoning, genomics analysis, biochemistry knowledge, and scienti

Hierarchical DLO Routing with Reinforcement Learning and In-Context Vision-language Models

AgentsDGX agent

arXiv:2510.19268v2 Announce Type: replace-cross Abstract: Long-horizon routing tasks of deformable linear objects (DLOs), such as cables and ropes, are common in industrial assembly lines and everyday

New in LangSmith Evaluation: ✅ Evaluator template library ✅ Reusable evaluators Everything you need to know → https://www.langchain.com/blog…

AgentsDGX agent

New in LangSmith Evaluation: ✅ Evaluator template library ✅ Reusable evaluators Everything you need to know → https://www.langchain.com/blog/reusable-langsmith-evaluator-templates?utm_source=x&utm_med

Online Navigation Planning for Long-term Autonomous Operation of Underwater Gliders

AgentsDGX agent

arXiv:2602.19315v2 Announce Type: replace Abstract: Underwater glider robots have become indispensable for ocean sampling, yet fully autonomous long-term operation remains rare in practice. Although s

OpenAI’s big Codex update is a direct shot at Claude Code

Model ReleasesDGX agent

OpenAI is beefing up its agentic coding and development system, Codex, with a suite of updates that let it use your computer, generate images, and remember from past experiences. The package of update

Vision-and-Language Navigation for UAVs: Progress, Challenges, and a Research Roadmap

AgentsDGX agent

arXiv:2604.13654v1 Announce Type: new Abstract: Vision-and-Language Navigation for Unmanned Aerial Vehicles (UAV-VLN) represents a pivotal challenge in embodied artificial intelligence, focused on ena

We’ve also added support for 90+ plugins in Codex, giving it more ways to gather context and take action across the tools you already use fo…

Model ReleasesDGX agent

We’ve also added support for 90+ plugins in Codex, giving it more ways to gather context and take action across the tools you already use for docs, project management, code review, creative work, depl

15 Apr 2026

D-BDM: A Direct and Efficient Boundary-Based Occupancy Grid Mapping Framework for LiDARs

AgentsDGX agent

arXiv:2604.12436v1 Announce Type: new Abstract: Efficient and scalable 3D occupancy mapping is essential for autonomous robot applications in unknown environments. However, traditional occupancy grid

Devin runs in its own VM in the cloud. You plan locally, delegate to Devin with one click, and keep coding (or close your laptop). Included …

ToolsDGX agent

Devin runs in its own VM in the cloud. You plan locally, delegate to Devin with one click, and keep coding (or close your laptop). Included with every Windsurf plan. Access to Devin in Windsurf is rol

@lmsysorg @sgl_project @vllm_project @CoreWeave @nebiusai @nscale @togethercompute @Togethercompute enables AI labs and enterprises to run f…

HardwareDGX agent

@lmsysorg @sgl_project @vllm_project @CoreWeave @nebiusai @nscale @togethercompute @Togethercompute enables AI labs and enterprises to run frontier models and agents on the NVIDIA Blackwell platform—g

Long-Horizon Plan Execution in Large Tool Spaces through Entropy-Guided Branching

Model ReleasesDGX agent

arXiv:2604.12126v1 Announce Type: new Abstract: Large Language Models (LLMs) have significantly advanced tool-augmented agents, enabling autonomous reasoning via API interactions. However, executing m

Narrative-Driven Paper-to-Slide Generation via ArcDeck

Model ReleasesDGX agent

arXiv:2604.11969v1 Announce Type: new Abstract: We introduce ArcDeck, a multi-agent framework that formulates paper-to-slide generation as a structured narrative reconstruction task. Unlike existing m

← Previous
1…179180181182183…300
Next →