AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,771 results
5 May 2026

Assistance Without Interruption: A Benchmark and LLM-based Framework for Non-Intrusive Human-Robot Assistance

Model ReleasesDGX agent

arXiv:2605.01368v1 Announce Type: new Abstract: Human-robot interaction (HRI) has long studied how agents and people coordinate to achieve shared goals. In this work, we formalize and benchmark the no

Breaking the Computational Barrier: Provably Efficient Actor-Critic for Low-Rank MDPs

SafetyDGX agent

arXiv:2605.01242v1 Announce Type: new Abstract: Reinforcement learning (RL) is a fundamental framework for sequential decision-making, in which an agent learns an optimal policy through interactions w

GRAVITY: Architecture-Agnostic Structured Anchoring for Long-Horizon Conversational Memory

ResearchDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.01688v1 Announce Type: new Abstract: Long-horizon conversational agents rely on memory systems with increasingly sophisticated retrieval mechanisms. However, retrieved fragments are typical

Hierarchical Federated Learning for Networked AI: From Communication Saving to Architecture-Aware Design

Local AiDGX agent

arXiv:2605.00931v1 Announce Type: new Abstract: Federated learning (FL) is fundamentally a distributed optimization problem executed by communicating agents with local data, local computation, and par

HyMem: Hybrid Memory Architecture with Dynamic Retrieval Scheduling

ResearchDGX agent

arXiv:2602.13933v2 Announce Type: replace Abstract: Large language model (LLM) agents demonstrate strong performance in short-text contexts but often underperform in extended dialogues due to ineffici

Multi-User Dueling Bandits: A Fair Approach using Nash Social Welfare

SafetyDGX agent

arXiv:2605.01961v1 Announce Type: new Abstract: Learning from human preference data is becoming a useful tool, from fine-tuning large language models to training reinforcement learning agents. However

Open models should compete on cost and specialization, not frontier benchmarks @natolambert puts it well: the right benchmark is savings in …

Model ReleasesDGX agent

Open models should compete on cost and specialization, not frontier benchmarks @natolambert puts it well: the right benchmark is savings in compute and time, especially for repetitive agent tasks deep

.@Redisinc is returning to Interrupt! Say hello to their team at the expo hall to learn about their fast memory layer for chatbots and AI ag…

TutorialsDGX agent

.@Redisinc is returning to Interrupt! Say hello to their team at the expo hall to learn about their fast memory layer for chatbots and AI agents as well as their ready-to-use tools for building AI app

Robust volatility updates for Hierarchical Gaussian Filtering

Model ReleasesDGX agent

arXiv:2605.00966v1 Announce Type: new Abstract: Hierarchical Gaussian Filtering (HGF) networks allow for efficient updating of posterior distributions (beliefs) about hidden states of an agent's envir

The Compliance Gap: Why AI Systems Promise to Follow Process Instructions but Don't

Model ReleasesDGX agent

arXiv:2605.01771v1 Announce Type: new Abstract: An auditor instructs an AI assistant: 'open each file individually using the Read tool -- no scripts, no agents.' The AI replies 'Yes' -- then issues a

Today, the CEO of @coinbase announced a 14% headcount reduction. Let's talk about why. He cites AI as a core reason for the shift, alongside…

SafetyDGX agent

Today, the CEO of @coinbase announced a 14% headcount reduction. Let's talk about why. He cites AI as a core reason for the shift, alongside market dynamics - though we don't know whether each of thos

When Correct Isn't Usable: Improving Structured Output Reliability in Small Language Models

Model ReleasesDGX agent

arXiv:2605.02363v1 Announce Type: new Abstract: Deployed language models must produce outputs that are both correct and format-compliant. We study this structured-output reliability gap using two math

4 May 2026

LLM-Oriented Information Retrieval: A Denoising-First Perspective

Model ReleasesDGX agent

arXiv:2605.00505v1 Announce Type: cross Abstract: Modern information retrieval (IR) is no longer consumed primarily by humans but increasingly by large language models (LLMs) via retrieval-augmented g

Scaling Federated Linear Contextual Bandits via Sketching

ApplicationsDGX agent

arXiv:2605.00500v1 Announce Type: new Abstract: In federated contextual linear bandits, high data dimensionality incurs prohibitive computation and communication costs: local agents perform O(d^3)-tim

ScreenParse: Moving Beyond Sparse Grounding with Complete Screen Parsing Supervision

Model ReleasesDGX agent

arXiv:2602.14276v2 Announce Type: replace Abstract: Modern computer-use agents (CUA) must perceive a screen as a structured state, what elements are visible, where they are, and what text they contain

Value Explicit Pretraining for Learning Transferable Representations

Model ReleasesDGX agent

arXiv:2312.12339v3 Announce Type: replace Abstract: Understanding visual inputs for a given task amidst varied changes is a key challenge posed by visual reinforcement learning agents. We propose exti

3 May 2026

that’s right.

SafetyDGX agent

that’s right. 😂 Finally, the great LeCun conversion arc is complete! “LLM agents = disaster” – said the guy who once defended Galactica like it was the Second Coming. Gary, you didn’t change… the time

2 May 2026

Grok 4.3 - excellent intelligence per unit cost

Model ReleasesDGX agent

Grok 4.3 - excellent intelligence per unit cost xAI has launched Grok 4.3, achieving 53 on the Artificial Analysis Intelligence Index with improved agentic performance, ~40% lower input price, and ~60

1 May 2026

Attractor FCM

Local AiDGX agent

arXiv:2604.27947v1 Announce Type: cross Abstract: In this paper an attractor FCM is created, tested, and analyzed. This FCM is neither a hebbian based nor agentic, nor a hybrid; it rather is a gradien

AWS Transform now automates BI migration to Amazon Quick in days

IndustryDGX agent

In this post, we walk through the full journey, from setting up your migration workspace in AWS Transform to subscribing to partner agents through AWS Marketplace to unlocking Amazon Quick capabilitie

Cofounder 2 launches on May 4th.

Model ReleasesDGX agent

Yohei Nakajima announced the launch of Cofounder 2 on May 4th via X. Cofounder is an AI agent tool designed to assist with business and startup tasks. The announcement was shared on social media to in

D3-Gym: Constructing Real-World Verifiable Environments for Data-Driven Discovery

SafetyDGX agent

arXiv:2604.27977v1 Announce Type: new Abstract: Despite recent progress in language models and agents for scientific data-driven discovery, further advancing their capabilities is held back by the abs

Detecting is Easy, Adapting is Hard: Local Expert Growth for Visual Model-Based Reinforcement Learning under Distribution Shift

Local AiDGX agent

arXiv:2604.27411v1 Announce Type: new Abstract: Visual model-based reinforcement learning (MBRL) agents can perform well on the training distribution, but often break down once the test environment sh

EdgeFM: Efficient Edge Inference for Vision-Language Models

HardwareDGX agent

arXiv:2604.27476v1 Announce Type: new Abstract: Vision-language models (VLMs) have demonstrated strong applicability in edge industrial applications, yet their deployment remains severely constrained

Graph World Models: Concepts, Taxonomy, and Future Directions

TutorialsDGX agent

arXiv:2604.27895v1 Announce Type: new Abstract: As one of the mainstream models of artificial intelligence, world models allow agents to learn the representation of the environment for efficient predi

Learning to Forget: Continual Learning with Adaptive Weight Decay

Model ReleasesDGX agent

arXiv:2604.27063v1 Announce Type: new Abstract: Continual learning agents with finite capacity must balance acquiring new knowledge with retaining the old. This requires controlled forgetting of knowl

NashPG: A Policy Gradient Method with Iteratively Refined Regularization for Finding Nash Equilibria

Model ReleasesDGX agent

arXiv:2510.18183v2 Announce Type: replace Abstract: Finding Nash equilibria in two-player zero-sum imperfect-information games remains a central challenge in multi-agent reinforcement learning. Recent

Simulating Infant First-Person Sensorimotor Experience via Motion Retargeting from Babies to Humanoids

ResearchDGX agent

arXiv:2604.27583v1 Announce Type: cross Abstract: Motion retargeting from humans to human-like artificial agents is becoming increasingly important as humanoid robots grow more capable. However, most

What Suppresses Nash Equilibrium Play in Large Language Models? Mechanistic Evidence and Causal Control

Model ReleasesDGX agent

arXiv:2604.27167v1 Announce Type: cross Abstract: LLM agents are known to deviate from Nash equilibria in strategic interactions, but nobody has looked inside the model to understand why, or asked whe

30 Apr 2026

A Decision-Theoretic Formalisation of Steganography With Applications to LLM Monitoring

ResearchDGX agent

arXiv:2602.23163v3 Announce Type: replace Abstract: Large language models are beginning to show steganographic capabilities. Such capabilities could allow misaligned models to evade oversight mechanis

Beyond Screenshots: Evaluating VLMs' Understanding of UI Animations

ResearchDGX agent

arXiv:2604.26148v1 Announce Type: cross Abstract: AI agents operating on user interfaces must understand how interfaces communicate state and feedback to act reliably. As a core communicative modality

Distill-Belief: Closed-Loop Inverse Source Localization and Characterization in Physical Fields

Local AiDGX agent

arXiv:2604.26095v1 Announce Type: new Abstract: {Closed-loop inverse source localization and characterization (ISLC) requires a mobile agent to select measurements that localize sources and infer late

Inferix: A Block-Diffusion based Next-Generation Inference Engine for World Simulation

Model ReleasesDGX agent

arXiv:2511.20714v2 Announce Type: replace-cross Abstract: World models serve as core simulators for fields such as agentic AI, embodied AI, and gaming, capable of generating long, physically realistic

Join us Tue 5/5: #DeepSeek-V4's hybrid attention + sparse MoE reduces KV cache up to 90%, enabling 1M-token context. We'll cover why that ma…

Model ReleasesDGX agent

Join us Tue 5/5: #DeepSeek-V4's hybrid attention + sparse MoE reduces KV cache up to 90%, enabling 1M-token context. We'll cover why that makes it great for agentic workflows, what it took to serve at

LLM Psychosis: A Theoretical and Diagnostic Framework for Reality-Boundary Failures in Large Language Models

Model ReleasesDGX agent

arXiv:2604.25934v1 Announce Type: cross Abstract: The deployment of large language models (LLMs) as interactive agents has exposed a category of behavioral failure that prevailing terminology, princip

We’re continuing to improve the runtime, harness, and models powering Cursor Security Review for a strong out-of-the-box experience. Securit…

TutorialsDGX agent

We’re continuing to improve the runtime, harness, and models powering Cursor Security Review for a strong out-of-the-box experience. Security agents draw from your existing usage pool. Learn more: htt

We've partnered with @OpenAI to offer GPT-5.5 in Devin at 50% off through May 14 starting today.

Model ReleasesDGX agent

We've partnered with @OpenAI to offer GPT-5.5 in Devin at 50% off through May 14 starting today. GPT-5.5 is now available in Devin as an Agent Preview! GPT-5.5 has set a new bar for what's possible wi

29 Apr 2026

DeepSeek-V4 Pro now available on Together AI

Model ReleasesDGX agent

DeepSeek-V4 Pro is now available on Together AI with 512K context, controllable reasoning modes, and cached-input pricing for long-context reasoning workloads like code agents, document intelligence,

🚀 Introducing FlashQLA: high-performance linear attention kernels built on TileLang. ⚡ 2–3× forward speedup. 2× backward speedup. 💻 Purpos…

Model ReleasesDGX agent

🚀 Introducing FlashQLA: high-performance linear attention kernels built on TileLang. ⚡ 2–3× forward speedup. 2× backward speedup. 💻 Purpose-built for agentic AI on your personal devices. 💡Key insights

Salesforce introduces Agentforce Operations to automate outdated back-office tasks

IndustryDGX agent

Salesforce Inc. today launched Agentforce Operations, an artificial intelligence system designed to extend specialized AI agents into the back office to automate manual work. Workflow automation tools

Thinking About Thinking: Evaluating Reasoning in Post-Trained Language Models

SafetyDGX agent

arXiv:2510.16340v2 Announce Type: replace Abstract: Recent advances in post-training techniques have endowed Large Language Models (LLMs) with enhanced capabilities for tackling complex, logic-intensi

Yesterday, @xenovacom assembled Reachy Mini and this morning I created my first app for this new version in under an hour with Claude! Try i…

Model ReleasesDGX agent

Yesterday, @xenovacom assembled Reachy Mini and this morning I created my first app for this new version in under an hour with Claude! Try it on your Reachy mini (or in simulation): https://huggingfac

28 Apr 2026

A few notes on how to get started with building LLM Knowledge Bases. @karpathy popularized it but most people don't know where to start. Eve…

TutorialsDGX agent

A few notes on how to get started with building LLM Knowledge Bases. @karpathy popularized it but most people don't know where to start. Everyone should be creating LLM Wikis. Live session tomorrow. S

Benchmarking Source-Sensitive Reasoning in Turkish: Humans and LLMs under Evidential Trust Manipulation

Local AiDGX agent

arXiv:2604.24665v1 Announce Type: cross Abstract: This paper investigates whether source trustworthiness shapes Turkish evidential morphology and whether large language models (LLMs) track this sensit

Both are great models and neither wins everywhere. I use both Opus and 5.5 depending on the task. LangSmith Fleet lets you choose the model …

Model ReleasesDGX agent

Both are great models and neither wins everywhere. I use both Opus and 5.5 depending on the task. LangSmith Fleet lets you choose the model for each agent, so you can match it to the work https://www.

Domain-Filtered Knowledge Graphs from Sparse Autoencoder Features

Local AiDGX agent

arXiv:2604.23829v1 Announce Type: new Abstract: Sparse autoencoders (SAEs) extract millions of interpretable features from a language model, but flat feature inventories aren't very useful on their ow

Evaluating whether AI models would sabotage AI safety research

Model ReleasesDGX agent

arXiv:2604.24618v1 Announce Type: new Abstract: We evaluate the propensity of frontier models to sabotage or refuse to assist with safety research when deployed as AI research agents within a frontier

Grammar-Constrained Refinement of Safety Operational Rules Using Language in the Loop: What Could Go Wrong

SafetyDGX agent

arXiv:2604.23523v1 Announce Type: cross Abstract: Safety specifications in cyber-physical systems (CPS) capture the operational conditions the system must satisfy to operate safely within its intended

I'm so confused…

Model ReleasesDGX agent

I'm so confused… We're excited to partner with Google to offer Grounding With Exa inside of Gemini models! Using Exa's agent-first search, Gemini models can now access billions of websites, technical

InCoM: Intent-Driven Perception and Structured Coordination for Mobile Manipulation

SafetyDGX agent

arXiv:2602.23024v2 Announce Type: replace Abstract: Mobile manipulation is a fundamental capability for general-purpose robotic agents, requiring both coordinated control of the mobile base and manipu

IndustryAssetEQA: A Neurosymbolic Operational Intelligence System for Embodied Question Answering in Industrial Asset Maintenance

SafetyDGX agent

arXiv:2604.23446v1 Announce Type: new Abstract: Industrial maintenance environments increasingly rely on AI systems to assist operators in understanding asset behavior, diagnosing failures, and evalua

Intervention-Aware Multiscale Representation Learning from Imaging Phenomics and Perturbation Transcriptomics

SafetyDGX agent

arXiv:2604.22832v1 Announce Type: cross Abstract: Microscopy-based phenotypic profiling is scalable for drug discovery but lacks the mechanistic depth of transcriptomics, which remains costly and scar

Learning Selective LLM Autonomy from Copilot Feedback in Enterprise Customer Support Workflows

SafetyDGX agent

arXiv:2604.23855v1 Announce Type: new Abstract: We present a deployed system that automates end-to-end customer support workflows inside an enterprise Business Process Management (BPM) platform. The a

Let's talk document formatting. Bold. Italics. Superscripts. Strikethroughs. The visual cues humans rely on every time we read a doc, and on…

Model ReleasesDGX agent

Let's talk document formatting. Bold. Italics. Superscripts. Strikethroughs. The visual cues humans rely on every time we read a doc, and ones existing OCR benchmarks completely ignore. 😱'199' struck

LOCAL AI MODELS ARE CATCHING UP TO FRONTIER MODELS WAY FASTER THAN ANYONE EXPECTED this guy ran qwen 3.6 27B locally on a base macbook pro M…

Model ReleasesDGX agent

LOCAL AI MODELS ARE CATCHING UP TO FRONTIER MODELS WAY FASTER THAN ANYONE EXPECTED this guy ran qwen 3.6 27B locally on a base macbook pro M4 with 24GB of memory quantized and stripped of safety guard

RAT: RunAnyThing via Fully Automated Environment Configuration

Model ReleasesDGX agent

arXiv:2604.23190v1 Announce Type: cross Abstract: Automating repository-level software engineering tasks is a foundational challenge for autonomous code agents, largely due to the difficulty of config

SMP: Reusable Score-Matching Motion Priors for Physics-Based Character Control

SafetyDGX agent

arXiv:2512.03028v3 Announce Type: replace-cross Abstract: Data-driven motion priors that can guide agents toward producing naturalistic behaviors play a pivotal role in creating life-like virtual char

🆕 Today, we're releasing the public preview of Workflows, the orchestration layer for enterprise AI. 🌎 Enterprise teams have capable model…

ApplicationsDGX agent

🆕 Today, we're releasing the public preview of Workflows, the orchestration layer for enterprise AI. 🌎 Enterprise teams have capable models. What they don't have is a way to run them reliably in produ

When an LLM acts happy (“EUREKA!”) or sad (“I have failed…”), is that meaningless mimicry, or does it reflect something “real”? We don’t kno…

SafetyDGX agent

When an LLM acts happy (“EUREKA!”) or sad (“I have failed…”), is that meaningless mimicry, or does it reflect something “real”? We don’t know if LLMs are conscious. But they increasingly seem to exhib

27 Apr 2026

Context-Sensitive Abstractions for Reinforcement Learning with Parameterized Actions

TutorialsDGX agent

arXiv:2512.20831v2 Announce Type: replace Abstract: Real-world sequential decision-making often involves parameterized action spaces that require both, decisions regarding discrete actions and decisio

← Previous
1…235236237238239…297
Next →