AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,619Total entries
1Added by human
84,618Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,976 results
Agents

Position: Deployed Reinforcement Learning should be Continual

DGX agent

arXiv:2606.04029v1 Announce Type: cross Abstract: Reinforcement Learning (RL) has received increasing attention and adoption in real-world use cases. Most of these systems follow a train-then-fix para

agentsarxiv-cs-ai
4 Jun 2026
Agents

Stateful Visual Encoders for Vision-Language Models

X Post
Paper
YouTube
Reddit
GitHub
Clear filters
DGX agent

arXiv:2606.04433v1 Announce Type: cross Abstract: Vision-language models (VLMs) are increasingly used in multi-image, multi-turn agentic settings where decisions depend on visual changes. However, in

agentsarxiv-cs-cl
4 Jun 2026
Agents

Decentralized Stochastic Nonconvex Optimization under the (L_0,L_1)-Smoothness

DGX agent

arXiv:2509.08726v3 Announce Type: replace-cross Abstract: This paper focuses on the decentralized stochastic optimization problem f(mathbf{x})=frac{1}{m}sum_{i=1}^m f_i(mathbf{x}) over a connected net

agentsarxiv-cs-lg
3 Jun 2026
Model Releases

Dynamics of Cognitive Heterogeneity: Investigating Behavioral Biases in Multi-Stage Supply Chains with LLM-Based Simulation

DGX agent

arXiv:2604.17220v2 Announce Type: replace-cross Abstract: Modeling coordination among generative agents in complex multi-round decision-making presents a core challenge for AI and operations managemen

model-releasesarxiv-cs-ai
3 Jun 2026
Agents

latest iteration of middleware is rad: there was a subagent that we had that was taking some pretty wild trajectories and was costing wayyyy…

DGX agent

latest iteration of middleware is rad: there was a subagent that we had that was taking some pretty wild trajectories and was costing wayyyy too much We adapted a 60 line middleware from a different a

agentsharrison-chase--x
3 Jun 2026
Agents

microsoft MAI tech report is a gold mine, one of the most transparent for a model at this scale. this model uses zero synthetic data or dist…

DGX agent

microsoft MAI tech report is a gold mine, one of the most transparent for a model at this scale. this model uses zero synthetic data or distillation from previous models. this means reasoning, agentic

agentsswyx--x
3 Jun 2026
Agents

Minimax Optimal Strategy for Delayed Observations in Online Reinforcement Learning

DGX agent

arXiv:2603.03480v2 Announce Type: replace Abstract: We study reinforcement learning with delayed state observation, where the agent observes the current state after some random number of time steps. W

agentsarxiv-cs-lg
3 Jun 2026
Model Releases

Psi-Bench: Evaluating Persona-Sensitive Influencing in Persuasive Dialogues

DGX agent

arXiv:2606.02754v1 Announce Type: new Abstract: Personalization is a crucial capability of modern language agents. However, current research primarily positions personalized agents as passive responde

model-releasesarxiv-cs-lg
3 Jun 2026
Agents

Reinforcement Learning from Cross-domain Videos with Video Prediction Model

DGX agent

arXiv:2606.03201v1 Announce Type: cross Abstract: Reinforcement learning from expert videos across visually distinct domains is challenging due to the absence of reward signals and the presence of dom

agentsarxiv-cs-ai
3 Jun 2026
Agents

@swyx and I are curating the AI in GTM track at @aiDotEngineer on June 30. The thing every AI engineer must realize: GTM just became an engi…

DGX agent

@swyx and I are curating the AI in GTM track at @aiDotEngineer on June 30. The thing every AI engineer must realize: GTM just became an engineering problem. Outbound = agent design. Enrichment = retri

agentsswyx--x
3 Jun 2026
Agents

You shipped your app. Now what? Your app may look great, but if no one can find it, it stays invisible Publishing is only the beginning Meet…

DGX agent

You shipped your app. Now what? Your app may look great, but if no one can find it, it stays invisible Publishing is only the beginning Meet SEO Agent. It runs a scan for you and suggests fixes to hel

agentsreplit--x
3 Jun 2026
Model Releases

Auditing Asset-Specific Preferences in Financial Large Language Models: Evidence from Bitcoin Representations and Portfolio Allocation

DGX agent

arXiv:2606.02528v1 Announce Type: cross Abstract: Large language models now power robo-advisors and trading agents, yet whether they carry built-in biases toward specific assets is largely untested. W

model-releasesarxiv-cs-lg
2 Jun 2026
Agents

Automated Conjecture Resolution with Formal Verification

DGX agent

arXiv:2604.03789v2 Announce Type: replace-cross Abstract: Recent advances in large language models have significantly improved their ability to perform mathematical reasoning, extending from elementar

agentsarxiv-cs-ai
2 Jun 2026
Safety

Beyond Independent Manipulation: Individual Fairness-aware Strategic Classification with Peer Imitation

DGX agent

arXiv:2606.00827v1 Announce Type: cross Abstract: Strategic classification (SC) investigates scenarios where agents manipulate their features to obtain favorable decisions from predictive models. Exis

safetyarxiv-cs-ai
2 Jun 2026
Agents

Everyone talks about 1M context. The harder part is making 1M context actually usable. Serving MiniMax M3 required optimizing for long-conte…

DGX agent

Everyone talks about 1M context. The harder part is making 1M context actually usable. Serving MiniMax M3 required optimizing for long-context, multimodal, and agentic workloads simultaneously. Excite

agentstogether-ai--x
2 Jun 2026
Agents

Fiduciary grade AI sets the bar as Thomson Reuters and Snowflake bring governed intelligence to the professions

DGX agent

Professionals who carry personal liability for their decisions — lawyers, tax accountants, auditors — cannot afford AI that gets it wrong. As enterprises accelerate deployment of agentic systems, the

agentssiliconangle
2 Jun 2026
Model Releases

From Segments to Scenes: Temporal Understanding in Autonomous Driving via Vision-Language Model

DGX agent

arXiv:2512.05277v3 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) are increasingly deployed as the perception and reasoning backbone of autonomous agents acting in the wild, with

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

GenPT: Beyond Self-Report for Reliable LLM Psychometrics via Generative Projective Testing

DGX agent

arXiv:2606.00860v1 Announce Type: cross Abstract: Self-report questionnaires remain the prevailing tool for probing the psychological states of persona-conditioned agents (PC-Agents). However, classic

model-releasesarxiv-cs-ai
2 Jun 2026
Agents

micropython-wasm 0.1a1

DGX agent

micropython-wasm 0.1a1 is a Python library for running a MicroPython sandbox using WebAssembly . This alpha release includes fixes for limitations discovered while building datasette-agent-micropython

agentssimon-willison
2 Jun 2026
Model Releases

Model-Native Computing Architecture: Envisioning Future System Architecture Through the Lens of Computer Architecture

DGX agent

arXiv:2606.00288v1 Announce Type: new Abstract: Large language models are undergoing a transition from model technology to system technology. As developers use Codex, Claude Code, AutoGPT, and related

model-releasesarxiv-cs-ai
2 Jun 2026
Agents

Needles at Scale: LLM-Assisted Target Selection for Windows Vulnerability Research

DGX agent

arXiv:2606.01364v1 Announce Type: cross Abstract: The attack surface of a modern operating system is a haystack: thousands of signed binaries and millions of functions, almost none relevant to any giv

agentsarxiv-cs-ai
2 Jun 2026
Agents

PlatonicNav: Unveiling Semantic Correspondence in Navigation with Platonic Topological Maps

DGX agent

arXiv:2606.01788v1 Announce Type: new Abstract: Embodied visual navigation, where an agent perceives a complex environment and acts to reach a goal from raw sensory input, underpins a wide range of ap

agentsarxiv-cs-cv
2 Jun 2026
Agents

Symmetry-Aware 9D Pose Estimation with Sim(3)-Consistent Feature and Spherical Inception Convolution

DGX agent

arXiv:2606.02219v1 Announce Type: new Abstract: Object pose estimation is a fundamental problem for an agent system to perceive or manipulate objects in images or videos. However, current instance-lev

agentsarxiv-cs-cv
2 Jun 2026
Model Releases

Task diversity produces systematic transfer but inhibits continual reinforcement learning

DGX agent

arXiv:2606.00880v1 Announce Type: cross Abstract: Continual reinforcement learning aims to produce agents that learn not only to improve at their current tasks but also to adapt as task distributions

model-releasesarxiv-cs-ai
2 Jun 2026
Agents

Towards Sparse Video Understanding and Reasoning

DGX agent

arXiv:2602.13602v2 Announce Type: replace Abstract: We present revise (nderline{Re}asoning with nderline{Vi}deo nderline{S}parsity), a multi-round agent for video question answering (VQA). Instead of

agentsarxiv-cs-cv
2 Jun 2026
Agents

Upwind integrates runtime cloud security with Cisco Cloud Control

DGX agent

Cloud security startup Upwind Security Inc. today announced it has integrated its runtime security platform with Cisco Cloud Control, the unified platform for agentic information technology operations

agentssiliconangle
2 Jun 2026
Agents

We Parse PDFs We spent 7 figures to put this on billboards throughout SF. I thought long and hard about putting something more creative and …

DGX agent

We Parse PDFs We spent 7 figures to put this on billboards throughout SF. I thought long and hard about putting something more creative and whimsical. But then you wouldn’t know what we do. AI agents

agentsjerry-liu--x
2 Jun 2026
Agents

A Tight Theory of Error Feedback Algorithms in Distributed Optimization

DGX agent

arXiv:2605.31594v1 Announce Type: new Abstract: Communication costs are a major bottleneck in distributed learning and first-order optimization. A common approach to alleviate this issue is to compres

agentsarxiv-cs-lg
1 Jun 2026
Agents

Answer-Set-Programming-based Abstractions for Reinforcement Learning

DGX agent

arXiv:2605.31444v1 Announce Type: new Abstract: Reinforcement Learning (RL) enables autonomous agents to learn policies from experience, but realistic problems often involve enormous state spaces, mak

agentsarxiv-cs-ai
1 Jun 2026
Agents

great look into how Rippling built RipplingAI

DGX agent

great look into how Rippling built RipplingAI .@Rippling AI runs on Deep Agents and LangSmith. Here’s how they shipped to millions of users in 6 months. https://www.langchain.com/blog/how-rippling-wen

agentsharrison-chase--x
1 Jun 2026
Agents

If LLMs Have Human-Like Attributes, Then So Does Age of Empires II

DGX agent

arXiv:2605.31514v1 Announce Type: cross Abstract: Much research has been carried out on large language models (LLMs) and LLM-powered agentic workflows. However, many works within the field state emerg

agentsarxiv-cs-ai
1 Jun 2026
Agents

Learning to Perceive the World Through Control: Empowerment-Based Representation Learning

DGX agent

arXiv:2605.30656v1 Announce Type: new Abstract: In many practical reinforcement learning environments, observations are far higher-dimensional than the variables that matter for control. In this work,

agentsarxiv-cs-lg
1 Jun 2026
Agents

MiniMax-M3 will by arrive on HuggingFace openweight at next week!

DGX agent

MiniMax-M3 will by arrive on HuggingFace openweight at next week! Introducing MiniMax M3: The First Open-Weights Model to Combine Three Frontier Capabilities - Coding & Agentic Frontier: 59.0% SWE-Ben

agentsclem-delangue--x
1 Jun 2026
Agents

Open models!

DGX agent

Open models! Introducing MiniMax M3: The First Open-Weights Model to Combine Three Frontier Capabilities - Coding & Agentic Frontier: 59.0% SWE-Bench Pro, 66.0% Terminal Bench 2.1, 34.8% SWE-fficiency

agentsharrison-chase--x
1 Jun 2026
Agents

SpecDB: LLM-Generated Customized Databases via Feature-Oriented Decomposition

DGX agent

arXiv:2605.31097v1 Announce Type: cross Abstract: Mainstream relational databases ship a uniform feature set across deployments, although individual workloads exercise only a fraction of the available

agentsarxiv-cs-ai
1 Jun 2026
Model Releases

What’s new in Microsoft Foundry | May 2026

DGX agent

May ships trace-based evaluation for any agent on any cloud, Grok 4.3 and DeepSeek V4 in the model catalog, GPT-5 Reinforcement Fine-Tuning at gated GA, three Microsoft Research on-device agent models

model-releasesmicrosoft-foundry
31 May 2026
Agents

1/10 - PDF parsing at browser speed Jerry Liu showed LiteParse v2 using Rust to WebAssembly for sub-second extraction of messy PDFs, without…

DGX agent

1/10 - PDF parsing at browser speed Jerry Liu showed LiteParse v2 using Rust to WebAssembly for sub-second extraction of messy PDFs, without calling a model. It can sit as a default step in agent and

agentsjerry-liu--x
30 May 2026
Agents

The secret to LiteParse lies in the grid projection algorithm. We project a complex page layout with text and tables into well-structured te…

DGX agent

The secret to LiteParse lies in the grid projection algorithm. We project a complex page layout with text and tables into well-structured text, that humans can read and agents can understanding. This

agentsjerry-liu--x
30 May 2026
Agents

BitTP: The Lightweight Trajectory Prediction Model with BitLLM for Edge-Devices

DGX agent

arXiv:2605.29705v1 Announce Type: new Abstract: Trajectory prediction is a fundamental task for autonomous systems, requiring complex reasoning about multi-agent interactions and intents. Large langua

agentsarxiv-cs-ai
29 May 2026
Model Releases

Gram: Assessing sabotage propensities via automated alignment auditing

DGX agent

arXiv:2605.30322v1 Announce Type: cross Abstract: We introduce Gram, an automated alignment auditing framework to assess the propensity of AI agents to engage in sabotage. We evaluate Gemini models ac

model-releasesarxiv-cs-ai
29 May 2026
Agents

Honeyval: A Comprehensive Evaluation Framework for LLM-powered HTTP Honeypots

DGX agent

arXiv:2605.29963v1 Announce Type: cross Abstract: Honeypots are decoy systems mimicking real system components designed to defend against cyber attacks. Recently, LLMs increasingly serve as simulation

agentsarxiv-cs-ai
29 May 2026
Agents

Quality went up alongside output. Even with more PRs shipping, total incidents dropped 5%. They built security guardrails and quality standa…

DGX agent

Quality went up alongside output. Even with more PRs shipping, total incidents dropped 5%. They built security guardrails and quality standards into the agentic workflow itself. Productivity vs qualit

agentsboris-cherny--x
29 May 2026
Agents

The hand-wringing over token usage is real, but I don’t see any organization that has adopted AI retreating from use in coding or even consi…

DGX agent

The hand-wringing over token usage is real, but I don’t see any organization that has adopted AI retreating from use in coding or even considering it. We are a few months into agentic coding, and comp

agentsethan-mollick--x
29 May 2026
Agents

The teams seeing the biggest wins from AI are completely changing how they work, not speeding up what they already do. What steps can you de…

DGX agent

The teams seeing the biggest wins from AI are completely changing how they work, not speeding up what they already do. What steps can you delete, what handoffs go away, what can an agent just own end

agentsboris-cherny--x
29 May 2026
Model Releases

AsyncTool: Evaluating the Asynchronous Function Calling Capability under Multi-Task Scenarios

DGX agent

arXiv:2605.27995v1 Announce Type: new Abstract: Large language model (LLM)-based agents have shown strong capabilities in using external tools to solve complex tasks. However, existing evaluations oft

model-releasesarxiv-cs-ai
28 May 2026
Agents

AtomComposer: Discovering Chemical Space from First Principles with Reinforcement Learning

DGX agent

arXiv:2605.28287v1 Announce Type: new Abstract: Discovering novel stable molecules without training data remains a grand scientific challenge. Current molecular generative models are trained on large,

agentsarxiv-cs-lg
28 May 2026
Agents

DeltaMCP: Incremental Regeneration via Spec-Aware Transformation for MCP servers

DGX agent

arXiv:2605.28148v1 Announce Type: cross Abstract: The rapid development of LLMs coupled with the introduction of Model Context Protocol (MCP) has revolutionized how intelligent agents interact with AP

agentsarxiv-cs-ai
28 May 2026
Model Releases

Do LLMs Favor Their Providers? Measuring Vertical Integration Bias in Code Generation

DGX agent

arXiv:2605.28515v1 Announce Type: cross Abstract: Large Language Models (LLMs) have become an integral part of software development, especially with the advent of agentic capabilities. Yet, many front

model-releasesarxiv-cs-ai
28 May 2026
← Previous
1…210211212213214…375
Next →