AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,959 results
2 Jun 2026

PlatonicNav: Unveiling Semantic Correspondence in Navigation with Platonic Topological Maps

AgentsDGX agent

arXiv:2606.01788v1 Announce Type: new Abstract: Embodied visual navigation, where an agent perceives a complex environment and acts to reach a goal from raw sensory input, underpins a wide range of ap

Symmetry-Aware 9D Pose Estimation with Sim(3)-Consistent Feature and Spherical Inception Convolution

AgentsDGX agent

arXiv:2606.02219v1 Announce Type: new Abstract: Object pose estimation is a fundamental problem for an agent system to perceive or manipulate objects in images or videos. However, current instance-lev

Task diversity produces systematic transfer but inhibits continual reinforcement learning

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2606.00880v1 Announce Type: cross Abstract: Continual reinforcement learning aims to produce agents that learn not only to improve at their current tasks but also to adapt as task distributions

Towards Sparse Video Understanding and Reasoning

AgentsDGX agent

arXiv:2602.13602v2 Announce Type: replace Abstract: We present revise (nderline{Re}asoning with nderline{Vi}deo nderline{S}parsity), a multi-round agent for video question answering (VQA). Instead of

Upwind integrates runtime cloud security with Cisco Cloud Control

AgentsDGX agent

Cloud security startup Upwind Security Inc. today announced it has integrated its runtime security platform with Cisco Cloud Control, the unified platform for agentic information technology operations

We Parse PDFs We spent 7 figures to put this on billboards throughout SF. I thought long and hard about putting something more creative and …

AgentsDGX agent

We Parse PDFs We spent 7 figures to put this on billboards throughout SF. I thought long and hard about putting something more creative and whimsical. But then you wouldn’t know what we do. AI agents

1 Jun 2026

A Tight Theory of Error Feedback Algorithms in Distributed Optimization

AgentsDGX agent

arXiv:2605.31594v1 Announce Type: new Abstract: Communication costs are a major bottleneck in distributed learning and first-order optimization. A common approach to alleviate this issue is to compres

Answer-Set-Programming-based Abstractions for Reinforcement Learning

AgentsDGX agent

arXiv:2605.31444v1 Announce Type: new Abstract: Reinforcement Learning (RL) enables autonomous agents to learn policies from experience, but realistic problems often involve enormous state spaces, mak

great look into how Rippling built RipplingAI

AgentsDGX agent

great look into how Rippling built RipplingAI .@Rippling AI runs on Deep Agents and LangSmith. Here’s how they shipped to millions of users in 6 months. https://www.langchain.com/blog/how-rippling-wen

If LLMs Have Human-Like Attributes, Then So Does Age of Empires II

AgentsDGX agent

arXiv:2605.31514v1 Announce Type: cross Abstract: Much research has been carried out on large language models (LLMs) and LLM-powered agentic workflows. However, many works within the field state emerg

Learning to Perceive the World Through Control: Empowerment-Based Representation Learning

AgentsDGX agent

arXiv:2605.30656v1 Announce Type: new Abstract: In many practical reinforcement learning environments, observations are far higher-dimensional than the variables that matter for control. In this work,

MiniMax-M3 will by arrive on HuggingFace openweight at next week!

AgentsDGX agent

MiniMax-M3 will by arrive on HuggingFace openweight at next week! Introducing MiniMax M3: The First Open-Weights Model to Combine Three Frontier Capabilities - Coding & Agentic Frontier: 59.0% SWE-Ben

Open models!

AgentsDGX agent

Open models! Introducing MiniMax M3: The First Open-Weights Model to Combine Three Frontier Capabilities - Coding & Agentic Frontier: 59.0% SWE-Bench Pro, 66.0% Terminal Bench 2.1, 34.8% SWE-fficiency

SpecDB: LLM-Generated Customized Databases via Feature-Oriented Decomposition

AgentsDGX agent

arXiv:2605.31097v1 Announce Type: cross Abstract: Mainstream relational databases ship a uniform feature set across deployments, although individual workloads exercise only a fraction of the available

31 May 2026

What’s new in Microsoft Foundry | May 2026

Model ReleasesDGX agent

May ships trace-based evaluation for any agent on any cloud, Grok 4.3 and DeepSeek V4 in the model catalog, GPT-5 Reinforcement Fine-Tuning at gated GA, three Microsoft Research on-device agent models

30 May 2026

1/10 - PDF parsing at browser speed Jerry Liu showed LiteParse v2 using Rust to WebAssembly for sub-second extraction of messy PDFs, without…

AgentsDGX agent

1/10 - PDF parsing at browser speed Jerry Liu showed LiteParse v2 using Rust to WebAssembly for sub-second extraction of messy PDFs, without calling a model. It can sit as a default step in agent and

The secret to LiteParse lies in the grid projection algorithm. We project a complex page layout with text and tables into well-structured te…

AgentsDGX agent

The secret to LiteParse lies in the grid projection algorithm. We project a complex page layout with text and tables into well-structured text, that humans can read and agents can understanding. This

29 May 2026

BitTP: The Lightweight Trajectory Prediction Model with BitLLM for Edge-Devices

AgentsDGX agent

arXiv:2605.29705v1 Announce Type: new Abstract: Trajectory prediction is a fundamental task for autonomous systems, requiring complex reasoning about multi-agent interactions and intents. Large langua

Gram: Assessing sabotage propensities via automated alignment auditing

Model ReleasesDGX agent

arXiv:2605.30322v1 Announce Type: cross Abstract: We introduce Gram, an automated alignment auditing framework to assess the propensity of AI agents to engage in sabotage. We evaluate Gemini models ac

Honeyval: A Comprehensive Evaluation Framework for LLM-powered HTTP Honeypots

AgentsDGX agent

arXiv:2605.29963v1 Announce Type: cross Abstract: Honeypots are decoy systems mimicking real system components designed to defend against cyber attacks. Recently, LLMs increasingly serve as simulation

Quality went up alongside output. Even with more PRs shipping, total incidents dropped 5%. They built security guardrails and quality standa…

AgentsDGX agent

Quality went up alongside output. Even with more PRs shipping, total incidents dropped 5%. They built security guardrails and quality standards into the agentic workflow itself. Productivity vs qualit

The hand-wringing over token usage is real, but I don’t see any organization that has adopted AI retreating from use in coding or even consi…

AgentsDGX agent

The hand-wringing over token usage is real, but I don’t see any organization that has adopted AI retreating from use in coding or even considering it. We are a few months into agentic coding, and comp

The teams seeing the biggest wins from AI are completely changing how they work, not speeding up what they already do. What steps can you de…

AgentsDGX agent

The teams seeing the biggest wins from AI are completely changing how they work, not speeding up what they already do. What steps can you delete, what handoffs go away, what can an agent just own end

28 May 2026

AsyncTool: Evaluating the Asynchronous Function Calling Capability under Multi-Task Scenarios

Model ReleasesDGX agent

arXiv:2605.27995v1 Announce Type: new Abstract: Large language model (LLM)-based agents have shown strong capabilities in using external tools to solve complex tasks. However, existing evaluations oft

AtomComposer: Discovering Chemical Space from First Principles with Reinforcement Learning

AgentsDGX agent

arXiv:2605.28287v1 Announce Type: new Abstract: Discovering novel stable molecules without training data remains a grand scientific challenge. Current molecular generative models are trained on large,

DeltaMCP: Incremental Regeneration via Spec-Aware Transformation for MCP servers

AgentsDGX agent

arXiv:2605.28148v1 Announce Type: cross Abstract: The rapid development of LLMs coupled with the introduction of Model Context Protocol (MCP) has revolutionized how intelligent agents interact with AP

Do LLMs Favor Their Providers? Measuring Vertical Integration Bias in Code Generation

Model ReleasesDGX agent

arXiv:2605.28515v1 Announce Type: cross Abstract: Large Language Models (LLMs) have become an integral part of software development, especially with the advent of agentic capabilities. Yet, many front

Grimlock: Guarding High-Agency Systems with eBPF and Attested Channels

SafetyDGX agent

arXiv:2605.27488v1 Announce Type: cross Abstract: Agentic systems increasingly run user-authored orchestration code that invokes tools, spawns subtasks, and delegates work across machines and clouds.

One of the best ways to learn what LangSmith Engine is capable of is to talk to the team that built it. Join @bentannyhill for a live sessio…

AgentsDGX agent

One of the best ways to learn what LangSmith Engine is capable of is to talk to the team that built it. Join @bentannyhill for a live session on June 11th and see how your team can automate the agent

Safe In-Context Reinforcement Learning

Model ReleasesDGX agent

arXiv:2509.25582v3 Announce Type: replace Abstract: In-context reinforcement learning (ICRL) is an emerging RL paradigm where an agent, after pretraining, can adapt to out-of-distribution test tasks w

Self-Improving Language Models with Bidirectional Evolutionary Search

AgentsDGX agent

arXiv:2605.28814v1 Announce Type: new Abstract: Search has been proposed as an effective method for self-improving language models and agentic systems, both for post-training sample generation and for

The best design work doesn't happen in a chat box. You need space to explore ideas, create variants, and iterate Meet the new Replit Canvas …

AgentsDGX agent

The best design work doesn't happen in a chat box. You need space to explore ideas, create variants, and iterate Meet the new Replit Canvas Your agentic design tool to build beautiful websites, apps,

The Optimal Sample Complexity of Linear Contracts

AgentsDGX agent

arXiv:2601.01496v2 Announce Type: replace-cross Abstract: In this paper, we settle the problem of learning optimal linear contracts from data in the offline setting, where agent types are drawn from a

Writer helps solve brand consistency for enterprise marketing at scale

AgentsDGX agent

Writer Inc., an enterprise artificial intelligence agent platform used by leading enterprise brands to deliver their voice, today announced new infrastructure aimed at enforcing style, terminology and

27 May 2026

Can Retrieval Heads See Images? Multimodal Retrieval Heads in Long-Context Vision-Language Models

AgentsDGX agent

arXiv:2605.27243v1 Announce Type: new Abstract: Large vision-language models increasingly rely on long-context modeling to reason over documents, hour-level videos, and long-horizon agent trajectories

completing the loop is paramount most unrecognized step in the development lifecycle is iteration on the process itself

AgentsDGX agent

completing the loop is paramount most unrecognized step in the development lifecycle is iteration on the process itself everyone is talking about self-optimizing loops in software & agents. but what d

Excellent talk on engine The interest and adoption of engine has been far greater than we anticipated If you don’t know what it is, this is …

AgentsDGX agent

Excellent talk on engine The interest and adoption of engine has been far greater than we anticipated If you don’t know what it is, this is a great starting point Improving your agent has been a manua

Real-Time Progress Prediction in Reasoning Language Models

AgentsDGX agent

arXiv:2506.23274v4 Announce Type: replace-cross Abstract: Recent reasoning language models, particularly those that employ long latent chains of thought, achieve strong performance on complex agentic

Self-signals Driven Multi-LLM Debate for Efficient and Accurate Reasoning

AgentsDGX agent

arXiv:2510.06843v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have exhibited impressive capabilities across diverse application domains. Recent work has explored Multi-LLM Age

The Sensation Modulating Network:Haltability as the architectural ground for object-directed phenomenology

AgentsDGX agent

arXiv:2605.26856v1 Announce Type: cross Abstract: Cognitive science remains split between cognitivism - which accounts for recursion and language but cannot ground formal symbols in meaning - and 4E a

26 May 2026

AION: Next-Generation Tasks and Practical Harness for Time Series

AgentsDGX agent

arXiv:2605.25045v1 Announce Type: new Abstract: Time series research is moving beyond fixed forecasting benchmarks toward realistic tasks that combine prediction, contextual reasoning, tool use, and s

Claw-Anything: Benchmarking Always-On Personal Assistants with Broader Access to User's Digital World

Model ReleasesDGX agent

arXiv:2605.26086v1 Announce Type: new Abstract: Large language model agents are increasingly envisioned as always-on personal assistants with access to anything relevant in the user's digital world. Y

CoRe-Code: Collaborative Reinforcement Learning for Code Generation

Local AiDGX agent

arXiv:2605.24812v1 Announce Type: new Abstract: Large language models (LLMs) have achieved strong performance in code generation, but most methods rely on autoregressive decoding without global planni

Optimal Design for Multinomial Logit Model with Applications to Best Assortment Identification

AgentsDGX agent

arXiv:2605.25592v1 Announce Type: cross Abstract: We study optimal experimental design for multinomial logit (MNL) bandits, where an agent repeatedly selects a subset of K items from a ground set of s

PID-Guided Partial Alignment for Multimodal Decentralized Federated Learning

SafetyDGX agent

arXiv:2601.10012v2 Announce Type: replace Abstract: Multimodal decentralized federated learning (DFL) must support collaboration among agents that hold different modality subsets and often different m

SkillEvolBench: Benchmarking the Evolution from Episodic Experience to Procedural Skills

Model ReleasesDGX agent

arXiv:2605.24117v1 Announce Type: new Abstract: Large language model (LLM) agents accumulate rich episodic trajectories while solving real-world tasks, but it remains unclear whether such experience c

WorldGUI: An Interactive Benchmark for Desktop GUI Automation from Any Starting Point

Model ReleasesDGX agent

arXiv:2502.08047v5 Announce Type: replace Abstract: Recent progress in GUI agents has substantially improved visual grounding, yet robust planning remains challenging, particularly when the environmen

your bank knows what you did. it has no idea why. that context, the why behind every transaction, is the most valuable data in banking. and …

AgentsDGX agent

your bank knows what you did. it has no idea why. that context, the why behind every transaction, is the most valuable data in banking. and nobody's capturing it. this changes with agentic banking. Me

25 May 2026

Agentivism: a learning theory for the age of artificial intelligence

AgentsDGX agent

arXiv:2604.07813v2 Announce Type: replace Abstract: Learning theories have historically changed when the conditions of learning evolved. Generative and agentic AI create a new condition by allowing le

AI Assurance: A Comprehensive Testing Strategy for Enterprise AI Systems

AgentsDGX agent

arXiv:2605.23459v1 Announce Type: cross Abstract: Enterprise AI systems, built on large language models, retrieval pipelines and autonomous agents, introduce a class of risks that traditional software

Curriculum reinforcement learning with measurable task representation learning

AgentsDGX agent

arXiv:2605.23372v1 Announce Type: cross Abstract: In curriculum reinforcement learning (CRL), an agent incrementally accumulates knowledge over a sequence of tasks (i.e., a curriculum), and the learni

From Residuals to Reasons: LLM-Guided Mechanism Inference from Tabular Data

AgentsDGX agent

arXiv:2605.22897v1 Announce Type: new Abstract: A persistent challenge in machine learning for scientific applications is jointly achieving prediction and understanding. Statistical models excel on st

GT-HarmBench: Benchmarking AI Safety Risks Through the Lens of Game Theory

Model ReleasesDGX agent

arXiv:2602.12316v2 Announce Type: replace Abstract: Frontier AI systems are increasingly capable and deployed in high-stakes multi-agent environments. However, existing AI safety benchmarks largely ev

Query-Adaptive Semantic Chunking for Retrieval-Augmented Generation: A Dynamic Strategy with Contextual Window Expansion

AgentsDGX agent

arXiv:2605.22834v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) systems depend critically on document chunking quality for retrieving relevant context. Fixed chunking segments doc

Ralph Loops are powerful, but wrapping a naive loop in a shell script is a total token burner in production. Pinecone Principal Engineer Jen…

AgentsDGX agent

Ralph Loops are powerful, but wrapping a naive loop in a shell script is a total token burner in production. Pinecone Principal Engineer Jen Hamon breaks down why standard loops collapse: ❌ The Bug: P

Turning Adaptation into Assets: Cross-Domain Bridging for Online Vision-Language Navigation

AgentsDGX agent

arXiv:2605.23257v1 Announce Type: cross Abstract: Navigating under non-stationary environment shifts poses a critical challenge for a Vision-and-Language Navigation (VLN) agent deployed in the wild. Y

23 May 2026

Active Graph is the best, most 'correct' knowledge/context engine I've come across so far (and I've tried or at least researched most of the…

AgentsDGX agent

Active Graph is the best, most 'correct' knowledge/context engine I've come across so far (and I've tried or at least researched most of them.) babyagi has ~200 citations, but 0 papers... i just publi

22 May 2026

Brand new episode of Max Agency ⤵️

AgentsDGX agent

Brand new episode of Max Agency ⤵️ Great conversation on the Max Agency podcast with @cogent_security Co-Founder + CTO Geng Sng on building agents for autonomous cyber defense. Check out the full epis

Cohere Command A+ is now available in Microsoft Foundry as a Managed Compute offer. Cohere’s latest open-source model is built for enterpris…

AgentsDGX agent

Cohere Command A+ is now available in Microsoft Foundry as a Managed Compute offer. Cohere’s latest open-source model is built for enterprise-grade agentic AI workloads, bringing together reasoning, m

Couple days back @swyx posted a challenge: code a ~10M transformer in JAX/Flax/Optax, run it in free Colab, and train it on addition w/ your…

AgentsDGX agent

Couple days back @swyx posted a challenge: code a ~10M transformer in JAX/Flax/Optax, run it in free Colab, and train it on addition w/ your agent! I gave Codex the screenshot + /goal. It controlled C

← Previous
1…168169170171172…300
Next →