AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,919 results
Safety

Poly-EPO: Training Exploratory Reasoning Models

DGX agent

arXiv:2604.17654v3 Announce Type: replace Abstract: Exploration is a cornerstone of learning from experience: it enables agents to find solutions to complex problems, generalize to novel ones, and sca

safetyarxiv-cs-ai
6 May 2026
Model Releases

Safety and accuracy follow different scaling laws in clinical large language models

AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2605.04039v1 Announce Type: new Abstract: Clinical LLMs are often scaled by increasing model size, context length, retrieval complexity, or inference-time compute, with the implicit expectation

model-releasesarxiv-cs-cl
6 May 2026
Safety

The AI risk repository: A meta-review, database, and taxonomy of risks from artificial intelligence

DGX agent

arXiv:2408.12622v3 Announce Type: replace-cross Abstract: Artificial intelligence (AI) is reshaping society, from video generation to medical diagnosis, coding agents to autonomous vehicles. Yet resea

safetyarxiv-cs-lg
6 May 2026
Model Releases

Towards Understanding Specification Gaming in Reasoning Models

DGX agent

arXiv:2605.02269v1 Announce Type: new Abstract: Specification gaming is a critical failure mode of LLM agents. Despite this, there has been little systematic research into when it arises and what driv

model-releasesarxiv-cs-ai
6 May 2026
Tutorials

ValueBlindBench: Agreement-Gated Stress Testing of LLM-Judged Investment Rationales Before Returns Are Observable

DGX agent

arXiv:2604.25224v2 Announce Type: replace Abstract: LLM-based financial agents increasingly produce investment rationales before the outcomes needed to evaluate them are observable. This creates a del

tutorialsarxiv-cs-ai
6 May 2026
Model Releases

We're winding back our peak hours limit reduction and doubling 5 hour limits. Excited to partner with SpaceX to bring you more compute and w…

DGX agent

We're winding back our peak hours limit reduction and doubling 5 hour limits. Excited to partner with SpaceX to bring you more compute and we'll keep pushing to bring you the best coding agent in the

model-releasesthariq--x
6 May 2026
Safety

AFFormer: Adaptive Feature Fusion Transformer for V2X Cooperative Perception under Channel Impairments

DGX agent

arXiv:2605.01888v1 Announce Type: new Abstract: Accurate 3D object detection is essential for ensuring the safety of autonomous vehicles. Cooperative perception, which leverages vehicle-to-everything

safetyarxiv-cs-cv
5 May 2026
Model Releases

Assistance Without Interruption: A Benchmark and LLM-based Framework for Non-Intrusive Human-Robot Assistance

DGX agent

arXiv:2605.01368v1 Announce Type: new Abstract: Human-robot interaction (HRI) has long studied how agents and people coordinate to achieve shared goals. In this work, we formalize and benchmark the no

model-releasesarxiv-cs-ro
5 May 2026
Safety

Breaking the Computational Barrier: Provably Efficient Actor-Critic for Low-Rank MDPs

DGX agent

arXiv:2605.01242v1 Announce Type: new Abstract: Reinforcement learning (RL) is a fundamental framework for sequential decision-making, in which an agent learns an optimal policy through interactions w

safetyarxiv-cs-lg
5 May 2026
Research

GRAVITY: Architecture-Agnostic Structured Anchoring for Long-Horizon Conversational Memory

DGX agent

arXiv:2605.01688v1 Announce Type: new Abstract: Long-horizon conversational agents rely on memory systems with increasingly sophisticated retrieval mechanisms. However, retrieved fragments are typical

researcharxiv-cs-cl
5 May 2026
Local Ai

Hierarchical Federated Learning for Networked AI: From Communication Saving to Architecture-Aware Design

DGX agent

arXiv:2605.00931v1 Announce Type: new Abstract: Federated learning (FL) is fundamentally a distributed optimization problem executed by communicating agents with local data, local computation, and par

local-aiarxiv-cs-lg
5 May 2026
Research

HyMem: Hybrid Memory Architecture with Dynamic Retrieval Scheduling

DGX agent

arXiv:2602.13933v2 Announce Type: replace Abstract: Large language model (LLM) agents demonstrate strong performance in short-text contexts but often underperform in extended dialogues due to ineffici

researcharxiv-cs-ai
5 May 2026
Safety

Multi-User Dueling Bandits: A Fair Approach using Nash Social Welfare

DGX agent

arXiv:2605.01961v1 Announce Type: new Abstract: Learning from human preference data is becoming a useful tool, from fine-tuning large language models to training reinforcement learning agents. However

safetyarxiv-cs-lg
5 May 2026
Model Releases

Open models should compete on cost and specialization, not frontier benchmarks @natolambert puts it well: the right benchmark is savings in …

DGX agent

Open models should compete on cost and specialization, not frontier benchmarks @natolambert puts it well: the right benchmark is savings in compute and time, especially for repetitive agent tasks deep

model-releasesharrison-chase--x
5 May 2026
Tutorials

.@Redisinc is returning to Interrupt! Say hello to their team at the expo hall to learn about their fast memory layer for chatbots and AI ag…

DGX agent

.@Redisinc is returning to Interrupt! Say hello to their team at the expo hall to learn about their fast memory layer for chatbots and AI agents as well as their ready-to-use tools for building AI app

tutorialsharrison-chase--x
5 May 2026
Model Releases

Robust volatility updates for Hierarchical Gaussian Filtering

DGX agent

arXiv:2605.00966v1 Announce Type: new Abstract: Hierarchical Gaussian Filtering (HGF) networks allow for efficient updating of posterior distributions (beliefs) about hidden states of an agent's envir

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

The Compliance Gap: Why AI Systems Promise to Follow Process Instructions but Don't

DGX agent

arXiv:2605.01771v1 Announce Type: new Abstract: An auditor instructs an AI assistant: 'open each file individually using the Read tool -- no scripts, no agents.' The AI replies 'Yes' -- then issues a

model-releasesarxiv-cs-cl
5 May 2026
Safety

Today, the CEO of @coinbase announced a 14% headcount reduction. Let's talk about why. He cites AI as a core reason for the shift, alongside…

DGX agent

Today, the CEO of @coinbase announced a 14% headcount reduction. Let's talk about why. He cites AI as a core reason for the shift, alongside market dynamics - though we don't know whether each of thos

safetyallie-k--miller--x
5 May 2026
Model Releases

When Correct Isn't Usable: Improving Structured Output Reliability in Small Language Models

DGX agent

arXiv:2605.02363v1 Announce Type: new Abstract: Deployed language models must produce outputs that are both correct and format-compliant. We study this structured-output reliability gap using two math

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

LLM-Oriented Information Retrieval: A Denoising-First Perspective

DGX agent

arXiv:2605.00505v1 Announce Type: cross Abstract: Modern information retrieval (IR) is no longer consumed primarily by humans but increasingly by large language models (LLMs) via retrieval-augmented g

model-releasesarxiv-cs-cl
4 May 2026
Applications

Scaling Federated Linear Contextual Bandits via Sketching

DGX agent

arXiv:2605.00500v1 Announce Type: new Abstract: In federated contextual linear bandits, high data dimensionality incurs prohibitive computation and communication costs: local agents perform O(d^3)-tim

applicationsarxiv-cs-lg
4 May 2026
Model Releases

ScreenParse: Moving Beyond Sparse Grounding with Complete Screen Parsing Supervision

DGX agent

arXiv:2602.14276v2 Announce Type: replace Abstract: Modern computer-use agents (CUA) must perceive a screen as a structured state, what elements are visible, where they are, and what text they contain

model-releasesarxiv-cs-cv
4 May 2026
Model Releases

Value Explicit Pretraining for Learning Transferable Representations

DGX agent

arXiv:2312.12339v3 Announce Type: replace Abstract: Understanding visual inputs for a given task amidst varied changes is a key challenge posed by visual reinforcement learning agents. We propose exti

model-releasesarxiv-cs-lg
4 May 2026
Safety

that’s right.

DGX agent

that’s right. 😂 Finally, the great LeCun conversion arc is complete! “LLM agents = disaster” – said the guy who once defended Galactica like it was the Second Coming. Gary, you didn’t change… the time

safetygary-marcus--x
3 May 2026
Model Releases

Grok 4.3 - excellent intelligence per unit cost

DGX agent

Grok 4.3 - excellent intelligence per unit cost xAI has launched Grok 4.3, achieving 53 on the Artificial Analysis Intelligence Index with improved agentic performance, ~40% lower input price, and ~60

model-releaseselon-musk--x
2 May 2026
Local Ai

Attractor FCM

DGX agent

arXiv:2604.27947v1 Announce Type: cross Abstract: In this paper an attractor FCM is created, tested, and analyzed. This FCM is neither a hebbian based nor agentic, nor a hybrid; it rather is a gradien

local-aiarxiv-cs-ai
1 May 2026
Industry

AWS Transform now automates BI migration to Amazon Quick in days

DGX agent

In this post, we walk through the full journey, from setting up your migration workspace in AWS Transform to subscribing to partner agents through AWS Marketplace to unlocking Amazon Quick capabilitie

industryaws-ml-blog
1 May 2026
Model Releases

Cofounder 2 launches on May 4th.

DGX agent

Yohei Nakajima announced the launch of Cofounder 2 on May 4th via X. Cofounder is an AI agent tool designed to assist with business and startup tasks. The announcement was shared on social media to in

model-releasesyohei-nakajima--x
1 May 2026
Safety

D3-Gym: Constructing Real-World Verifiable Environments for Data-Driven Discovery

DGX agent

arXiv:2604.27977v1 Announce Type: new Abstract: Despite recent progress in language models and agents for scientific data-driven discovery, further advancing their capabilities is held back by the abs

safetyarxiv-cs-ai
1 May 2026
Local Ai

Detecting is Easy, Adapting is Hard: Local Expert Growth for Visual Model-Based Reinforcement Learning under Distribution Shift

DGX agent

arXiv:2604.27411v1 Announce Type: new Abstract: Visual model-based reinforcement learning (MBRL) agents can perform well on the training distribution, but often break down once the test environment sh

local-aiarxiv-cs-lg
1 May 2026
Hardware

EdgeFM: Efficient Edge Inference for Vision-Language Models

DGX agent

arXiv:2604.27476v1 Announce Type: new Abstract: Vision-language models (VLMs) have demonstrated strong applicability in edge industrial applications, yet their deployment remains severely constrained

hardwarearxiv-cs-cv
1 May 2026
Tutorials

Graph World Models: Concepts, Taxonomy, and Future Directions

DGX agent

arXiv:2604.27895v1 Announce Type: new Abstract: As one of the mainstream models of artificial intelligence, world models allow agents to learn the representation of the environment for efficient predi

tutorialsarxiv-cs-ai
1 May 2026
Model Releases

Learning to Forget: Continual Learning with Adaptive Weight Decay

DGX agent

arXiv:2604.27063v1 Announce Type: new Abstract: Continual learning agents with finite capacity must balance acquiring new knowledge with retaining the old. This requires controlled forgetting of knowl

model-releasesarxiv-cs-lg
1 May 2026
Model Releases

NashPG: A Policy Gradient Method with Iteratively Refined Regularization for Finding Nash Equilibria

DGX agent

arXiv:2510.18183v2 Announce Type: replace Abstract: Finding Nash equilibria in two-player zero-sum imperfect-information games remains a central challenge in multi-agent reinforcement learning. Recent

model-releasesarxiv-cs-lg
1 May 2026
Research

Simulating Infant First-Person Sensorimotor Experience via Motion Retargeting from Babies to Humanoids

DGX agent

arXiv:2604.27583v1 Announce Type: cross Abstract: Motion retargeting from humans to human-like artificial agents is becoming increasingly important as humanoid robots grow more capable. However, most

researcharxiv-cs-ro
1 May 2026
Model Releases

What Suppresses Nash Equilibrium Play in Large Language Models? Mechanistic Evidence and Causal Control

DGX agent

arXiv:2604.27167v1 Announce Type: cross Abstract: LLM agents are known to deviate from Nash equilibria in strategic interactions, but nobody has looked inside the model to understand why, or asked whe

model-releasesarxiv-cs-ai
1 May 2026
Research

A Decision-Theoretic Formalisation of Steganography With Applications to LLM Monitoring

DGX agent

arXiv:2602.23163v3 Announce Type: replace Abstract: Large language models are beginning to show steganographic capabilities. Such capabilities could allow misaligned models to evade oversight mechanis

researcharxiv-cs-ai
30 Apr 2026
Research

Beyond Screenshots: Evaluating VLMs' Understanding of UI Animations

DGX agent

arXiv:2604.26148v1 Announce Type: cross Abstract: AI agents operating on user interfaces must understand how interfaces communicate state and feedback to act reliably. As a core communicative modality

researcharxiv-cs-cl
30 Apr 2026
Local Ai

Distill-Belief: Closed-Loop Inverse Source Localization and Characterization in Physical Fields

DGX agent

arXiv:2604.26095v1 Announce Type: new Abstract: {Closed-loop inverse source localization and characterization (ISLC) requires a mobile agent to select measurements that localize sources and infer late

local-aiarxiv-cs-ai
30 Apr 2026
Model Releases

Inferix: A Block-Diffusion based Next-Generation Inference Engine for World Simulation

DGX agent

arXiv:2511.20714v2 Announce Type: replace-cross Abstract: World models serve as core simulators for fields such as agentic AI, embodied AI, and gaming, capable of generating long, physically realistic

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

Join us Tue 5/5: #DeepSeek-V4's hybrid attention + sparse MoE reduces KV cache up to 90%, enabling 1M-token context. We'll cover why that ma…

DGX agent

Join us Tue 5/5: #DeepSeek-V4's hybrid attention + sparse MoE reduces KV cache up to 90%, enabling 1M-token context. We'll cover why that makes it great for agentic workflows, what it took to serve at

model-releasestogether-ai--x
30 Apr 2026
Model Releases

LLM Psychosis: A Theoretical and Diagnostic Framework for Reality-Boundary Failures in Large Language Models

DGX agent

arXiv:2604.25934v1 Announce Type: cross Abstract: The deployment of large language models (LLMs) as interactive agents has exposed a category of behavioral failure that prevailing terminology, princip

model-releasesarxiv-cs-ai
30 Apr 2026
Tutorials

We’re continuing to improve the runtime, harness, and models powering Cursor Security Review for a strong out-of-the-box experience. Securit…

DGX agent

We’re continuing to improve the runtime, harness, and models powering Cursor Security Review for a strong out-of-the-box experience. Security agents draw from your existing usage pool. Learn more: htt

tutorialscursor--x
30 Apr 2026
Model Releases

We've partnered with @OpenAI to offer GPT-5.5 in Devin at 50% off through May 14 starting today.

DGX agent

We've partnered with @OpenAI to offer GPT-5.5 in Devin at 50% off through May 14 starting today. GPT-5.5 is now available in Devin as an Agent Preview! GPT-5.5 has set a new bar for what's possible wi

model-releasescognition-ai--x
30 Apr 2026
Model Releases

DeepSeek-V4 Pro now available on Together AI

DGX agent

DeepSeek-V4 Pro is now available on Together AI with 512K context, controllable reasoning modes, and cached-input pricing for long-context reasoning workloads like code agents, document intelligence,

model-releasestogether-ai-blog
29 Apr 2026
Model Releases

🚀 Introducing FlashQLA: high-performance linear attention kernels built on TileLang. ⚡ 2–3× forward speedup. 2× backward speedup. 💻 Purpos…

DGX agent

🚀 Introducing FlashQLA: high-performance linear attention kernels built on TileLang. ⚡ 2–3× forward speedup. 2× backward speedup. 💻 Purpose-built for agentic AI on your personal devices. 💡Key insights

model-releasesqwen--x
29 Apr 2026
Industry

Salesforce introduces Agentforce Operations to automate outdated back-office tasks

DGX agent

Salesforce Inc. today launched Agentforce Operations, an artificial intelligence system designed to extend specialized AI agents into the back office to automate manual work. Workflow automation tools

industrysiliconangle
29 Apr 2026
Safety

Thinking About Thinking: Evaluating Reasoning in Post-Trained Language Models

DGX agent

arXiv:2510.16340v2 Announce Type: replace Abstract: Recent advances in post-training techniques have endowed Large Language Models (LLMs) with enhanced capabilities for tackling complex, logic-intensi

safetyarxiv-cs-cl
29 Apr 2026
← Previous
1…296297298299300…374
Next →