AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,963 results
Industry

Cyera acquires Ryft to give enterprises traceable data access for AI agents

DGX agent

Artificial intelligence and data security company Cyera Ltd. announced today that it has acquired Ryft Data Inc., an Israeli startup with an automated data lake platform designed for enterprises deplo

industrysiliconangle
23 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Environmental Understanding Vision-Language Model for Embodied Agent

DGX agent

arXiv:2604.19839v1 Announce Type: cross Abstract: Vision-language models (VLMs) have shown strong perception and reasoning abilities for instruction-following embodied agents. However, despite these a

safetyarxiv-cs-ai
23 Apr 2026
Safety

ProMMSearchAgent: A Generalizable Multimodal Search Agent Trained with Process-Oriented Rewards

DGX agent

arXiv:2604.20486v1 Announce Type: new Abstract: Training multimodal agents via reinforcement learning for knowledge-intensive visual reasoning is fundamentally hindered by the extreme sparsity of outc

safetyarxiv-cs-cv
23 Apr 2026
Agents

zero shot Kimi K2.6, go try it out its a good model sir! this is @Kimi_Moonshot running on @togethercompute, @opencode harness prompt below…

DGX agent

zero shot Kimi K2.6, go try it out its a good model sir! this is @Kimi_Moonshot running on @togethercompute, @opencode harness prompt below👇 Media Introducing Kimi K2.6 from @Kimi_Moonshot, a multimod

agentstogether-ai--x
23 Apr 2026
Model Releases

Cyber Defense Benchmark: Agentic Threat Hunting Evaluation for LLMs in SecOps

DGX agent

arXiv:2604.19533v1 Announce Type: cross Abstract: We introduce the Cyber Defense Benchmark, a benchmark for measuring how well large language model (LLM) agents perform the core SOC analyst task of th

model-releasesarxiv-cs-ai
22 Apr 2026
Industry

Now Meta will track what employees do on their computers to train its AI agents

DGX agent

Meta employees' activity at work is now being used to train the company's AI agents. As reported by Reuters, Meta is installing a tool it calls Model Capability Initiative (MCI) on US-based employees'

industrythe-verge-ai
22 Apr 2026
Model Releases

SAGE-32B: Agentic Reasoning via Iterative Distillation

DGX agent

arXiv:2601.04237v2 Announce Type: replace Abstract: We demonstrate SAGE-32B, a 32 billion parameter language model that focuses on agentic reasoning and long range planning tasks. Unlike chat models t

model-releasesarxiv-cs-ai
22 Apr 2026
Applications

Boomi builds a role for agents and guardrails in the data-connected enterprise

DGX agent

Artificial intelligence and machine learning have been a central focus for Boomi LP over much of its 26 years as a data activation company. The organization began storing anonymized metadata from cust

applicationssiliconangle
21 Apr 2026
Agents

Generative midtended cognition and Artificial Intelligence. Thinging with thinging things

DGX agent

arXiv:2411.06812v2 Announce Type: replace-cross Abstract: This paper introduces the concept of ``generative midtended cognition'', exploring the integration of generative AI with human cognition. The

agentsarxiv-cs-lg
21 Apr 2026
Agents

Kimi is the current open-source SOTA on Artificial Analysis

DGX agent

Kimi is the current open-source SOTA on Artificial Analysis Moonshot’s Kimi K2.6 is the new leading open weights model. Kimi K2.6 lands at #4 on the Artificial Analysis Intelligence Index (54) behind

agentskimi-moonshot--x
21 Apr 2026
Model Releases

Let's talk parsing charts 📊📈. Last week we released ParseBench, the first document OCR benchmark for AI agents. New in ParseBench: ChartDa…

DGX agent

Let's talk parsing charts 📊📈. Last week we released ParseBench, the first document OCR benchmark for AI agents. New in ParseBench: ChartDataPointMatch. Most document look at a chart and OCR the captio

model-releasesjerry-liu--x
21 Apr 2026
Safety

OVOD-Agent: A Markov-Bandit Framework for Proactive Visual Reasoning and Self-Evolving Detection

DGX agent

arXiv:2511.21064v2 Announce Type: replace-cross Abstract: Open-Vocabulary Object Detection (OVOD) aims to enable detectors to generalize across categories by leveraging semantic information. Although

safetyarxiv-cs-cv
21 Apr 2026
Agents

SUSE and Vultr’s open cloud infrastructure push goes global

DGX agent

Global AI ambitions keep colliding with the same wall: cloud infrastructure that can’t keep up with performance demands without compromising where data lives and who controls it. As organizations look

agentssiliconangle
21 Apr 2026
Model Releases

GTA-2: Benchmarking General Tool Agents from Atomic Tool-Use to Open-Ended Workflows

DGX agent

arXiv:2604.15715v1 Announce Type: cross Abstract: The development of general-purpose agents requires a shift from executing simple instructions to completing complex, real-world productivity workflows

model-releasesarxiv-cs-ai
20 Apr 2026
Agents

InfoChess: A Game of Adversarial Inference and a Laboratory for Quantifiable Information Control

DGX agent

arXiv:2604.15373v1 Announce Type: cross Abstract: We propose InfoChess, a symmetric adversarial game that elevates competitive information acquisition to the primary objective. There is no piece captu

agentsarxiv-cs-ai
20 Apr 2026
Safety

Long-Term Memory for VLA-based Agents in Open-World Task Execution

DGX agent

arXiv:2604.15671v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have demonstrated significant potential for embodied decision-making; however, their application in complex chemical

safetyarxiv-cs-ro
20 Apr 2026
Model Releases

PolicyBank: Evolving Policy Understanding for LLM Agents

DGX agent

arXiv:2604.15505v1 Announce Type: cross Abstract: LLM agents operating under organizational policies must comply with authorization constraints typically specified in natural language. In practice, su

model-releasesarxiv-cs-ai
20 Apr 2026
Tools

We recently shipped quality-of-life improvements to the Cursor CLI to make working with agents in the terminal more delightful. Use /debug t…

DGX agent

We recently shipped quality-of-life improvements to the Cursor CLI to make working with agents in the terminal more delightful. Use /debug to find root causes and fix tricky bugs that are hard to repr

toolscursor--x
20 Apr 2026
Model Releases

Dive into Claude Code: The Design Space of Today's and Future AI Agent Systems

DGX agent

arXiv:2604.14228v1 Announce Type: cross Abstract: Claude Code is an agentic coding tool that can run shell commands, edit files, and call external services on behalf of the user. This study describes

model-releasesarxiv-cs-cl
17 Apr 2026
Safety

Enhancing LLM-based Search Agents via Contribution Weighted Group Relative Policy Optimization

DGX agent

arXiv:2604.14267v1 Announce Type: new Abstract: Search agents extend Large Language Models (LLMs) beyond static parametric knowledge by enabling access to up-to-date and long-tail information unavaila

safetyarxiv-cs-lg
17 Apr 2026
Tutorials

LLM agents loop, drift, and get stuck on hard reasoning tasks up to 30% of the time. Current fixes are either too blunt (hard step limits) o…

DGX agent

LLM agents loop, drift, and get stuck on hard reasoning tasks up to 30% of the time. Current fixes are either too blunt (hard step limits) or too expensive (LLM-as-judge adding 10-15% overhead per ste

tutorialsdair-ai--x
17 Apr 2026
Model Releases

SafeHarness: Lifecycle-Integrated Security Architecture for LLM-based Agent Deployment

DGX agent

arXiv:2604.13630v1 Announce Type: cross Abstract: The performance of large language model (LLM) agents depends critically on the execution harness, the system layer that orchestrates tool use, context

model-releasesarxiv-cs-ai
17 Apr 2026
Model Releases

VeruSAGE: A Study of Agent-Based Verification for Rust Systems

DGX agent

arXiv:2512.18436v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have shown impressive capability to understand and develop code. However, their capability to rigorously reason a

model-releasesarxiv-cs-ai
17 Apr 2026
Industry

AI labs are buying Slack, Jira, and email archives from defunct startups to build 'reinforcement learning gyms' and train AI agents in simulated workplaces (Anna Tong/Forbes)

DGX agent

Anna Tong / Forbes: AI labs are buying Slack, Jira, and email archives from defunct startups to build “reinforcement learning gyms” and train AI agents in simulated workplaces — Defunct startups are b

industrytechmeme
16 Apr 2026
Model Releases

LiveClawBench: Benchmarking LLM Agents on Complex, Real-World Assistant Tasks

DGX agent

arXiv:2604.13072v1 Announce Type: new Abstract: LLM-based agents are increasingly expected to handle real-world assistant tasks, yet existing benchmarks typically evaluate them under isolated sources

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

⚡ Meet Qwen3.6-35B-A3B:Now Open-Source!🚀🚀 A sparse MoE model, 35B total params, 3B active. Apache 2.0 license. 🔥 Agentic coding on par wi…

DGX agent

⚡ Meet Qwen3.6-35B-A3B:Now Open-Source!🚀🚀 A sparse MoE model, 35B total params, 3B active. Apache 2.0 license. 🔥 Agentic coding on par with models 10x its active size 📷 Strong multimodal perception an

model-releasesjeremy-howard--x
16 Apr 2026
Agents

Nous Research🤝fal Congratulations to @NousResearch on the Tool Gateway launch. Developer-first infrastructure is what will define the next …

DGX agent

Nous Research🤝fal Congratulations to @NousResearch on the Tool Gateway launch. Developer-first infrastructure is what will define the next phase of agentic applications. Tool Gateway is now live in No

agentsnous-research--x
16 Apr 2026
Model Releases

TREX: Automating LLM Fine-tuning via Agent-Driven Tree-based Exploration

DGX agent

arXiv:2604.14116v1 Announce Type: cross Abstract: While Large Language Models (LLMs) have empowered AI research agents to perform isolated scientific tasks, automating complex, real-world workflows, s

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

ViBES: A Conversational Agent with Behaviorally-Intelligent 3D Virtual Body

DGX agent

arXiv:2512.14234v2 Announce Type: replace Abstract: Human communication is inherently multimodal and social: words, prosody, and body language jointly carry intent. Yet most prior systems model human

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

AffectAgent: Collaborative Multi-Agent Reasoning for Retrieval-Augmented Multimodal Emotion Recognition

DGX agent

arXiv:2604.12735v1 Announce Type: new Abstract: LLM-based multimodal emotion recognition relies on static parametric memory and often hallucinates when interpreting nuanced affective states. In this p

model-releasesarxiv-cs-cv
15 Apr 2026
Hardware

Best part are retrieval and search details: >Agent queries differ from human queries & what good results means changes too >Parallel queries…

DGX agent

Best part are retrieval and search details: >Agent queries differ from human queries & what good results means changes too >Parallel queries and ranking are both tools to the same outcome >Top-K preci

hardwareswyx--x
15 Apr 2026
Agents

`deepagents deploy` now supports user scoped memory! add a user/ directory in your project so each user gets their own writable AGENTS.md, s…

DGX agent

`deepagents deploy` now supports user scoped memory! add a user/ directory in your project so each user gets their own writable AGENTS.md, seeded on first deploy and persisted across conversations. yo

agentsharrison-chase--x
15 Apr 2026
Model Releases

Drawing on Memory: Dual-Trace Encoding Improves Cross-Session Recall in LLM Agents

DGX agent

arXiv:2604.12948v1 Announce Type: new Abstract: LLM agents with persistent memory store information as flat factual records, providing little context for temporal reasoning, change tracking, or cross-

model-releasesarxiv-cs-ai
15 Apr 2026
Safety

Every Picture Tells a Dangerous Story: Memory-Augmented Multi-Agent Jailbreak Attacks on VLMs

DGX agent

arXiv:2604.12616v1 Announce Type: new Abstract: The rapid evolution of Vision-Language Models (VLMs) has catalyzed unprecedented capabilities in artificial intelligence; however, this continuous modal

safetyarxiv-cs-ai
15 Apr 2026
Model Releases

Frontier-Eng: Benchmarking Self-Evolving Agents on Real-World Engineering Tasks with Generative Optimization

DGX agent

arXiv:2604.12290v1 Announce Type: new Abstract: Current LLM agent benchmarks, which predominantly focus on binary pass/fail tasks such as code generation or search-based question answering, often negl

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

Is Gemma 4 26B MoE or 31B good as an MCP agent for coding with Xcode?

DGX agent

This r/ollama discussion explores the suitability of Google's Gemma 4 models — specifically the 26B Mixture of Experts (MoE) and 31B Dense variants — as MCP (Model Context Protocol) agents for coding

model-releasesr-ollama
15 Apr 2026
Safety

No More Stale Feedback: Co-Evolving Critics for Open-World Agent Learning

DGX agent

arXiv:2601.06794v2 Announce Type: replace Abstract: Critique-guided reinforcement learning (RL) has emerged as a powerful paradigm for training LLM agents by augmenting sparse outcome rewards with nat

safetyarxiv-cs-ai
15 Apr 2026
Model Releases

Policy-Invisible Violations in LLM-Based Agents

DGX agent

arXiv:2604.12177v1 Announce Type: new Abstract: LLM-based agents can execute actions that are syntactically valid, user-sanctioned, and semantically appropriate, yet still violate organizational polic

model-releasesarxiv-cs-ai
15 Apr 2026
Industry

Q&A with ElevenLabs co-founder Mati Staniszewski on how audio models work, the company's business model, the conversational Turing Test, voice agents, and more (John Collison/Cheeky Pint)

DGX agent

John Collison / Cheeky Pint: Q&A with ElevenLabs co-founder Mati Staniszewski on how audio models work, the company's business model, the conversational Turing Test, voice agents, and more — Mati Stan

industrytechmeme
15 Apr 2026
Model Releases

Spatial Atlas: Compute-Grounded Reasoning for Spatial-Aware Research Agent Benchmarks

DGX agent

arXiv:2604.12102v1 Announce Type: new Abstract: We introduce compute-grounded reasoning (CGR), a design paradigm for spatial-aware research agents in which every answerable sub-problem is resolved by

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

Thought-Retriever: Don't Just Retrieve Raw Data, Retrieve Thoughts for Memory-Augmented Agentic Systems

DGX agent

arXiv:2604.12231v1 Announce Type: new Abstract: Large language models (LLMs) have transformed AI research thanks to their powerful internal capabilities and knowledge. However, existing LLMs still fai

model-releasesarxiv-cs-cl
15 Apr 2026
Model Releases

Towards Long-horizon Agentic Multimodal Search

DGX agent

arXiv:2604.12890v1 Announce Type: cross Abstract: Multimodal deep search agents have shown great potential in solving complex tasks by iteratively collecting textual and visual evidence. However, mana

model-releasesarxiv-cs-ai
15 Apr 2026
Safety

A Dual-Positive Monotone Parameterization for Multi-Segment Bids and a Validity Assessment Framework for Reinforcement Learning Agent-based Simulation of Electricity Markets

DGX agent

arXiv:2604.10252v1 Announce Type: new Abstract: Reinforcement learning agent-based simulation (RL-ABS) has become an important tool for electricity market mechanism analysis and evaluation. In the mod

safetyarxiv-cs-ai
14 Apr 2026
Model Releases

From Translation to Superset: Benchmark-Driven Evolution of a Production AI Agent from Rust to Python

DGX agent

arXiv:2604.11518v1 Announce Type: cross Abstract: Cross-language migration of large software systems is a persistent engineering challenge, particularly when the source codebase evolves rapidly. We pr

model-releasesarxiv-cs-ai
14 Apr 2026
Safety

Hubble: An LLM-Driven Agentic Framework for Safe and Automated Alpha Factor Discovery

DGX agent

arXiv:2604.09601v1 Announce Type: new Abstract: Discovering predictive alpha factors in quantitative finance remains a formidable challenge due to the vast combinatorial search space and inherently lo

safetyarxiv-cs-ai
14 Apr 2026
Model Releases

I have a Macbook AIR M5 Base and I want to run an Agentic Coding program, similar to Claude Code or Codex. Besides the model, how do I do it? I've already tried with Ollama, VS Code, Opencode, and haven't been able to. (I'm not a developer, sorry)

DGX agent

This Reddit thread addresses a common challenge for non-developers trying to run a local agentic coding assistant on a MacBook Air M5: while tools like Ollama, VS Code, and OpenCode are the right piec

model-releasesr-ollama
14 Apr 2026
Agents

Improving Layout Representation Learning Across Inconsistently Annotated Datasets via Agentic Harmonization

DGX agent

arXiv:2604.11042v1 Announce Type: new Abstract: Fine-tuning object detection (OD) models on combined datasets assumes annotation compatibility, yet datasets often encode conflicting spatial definition

agentsarxiv-cs-cv
14 Apr 2026
Agents

Instructing LLMs to Negotiate using Reinforcement Learning with Verifiable Rewards

DGX agent

arXiv:2604.09855v1 Announce Type: new Abstract: The recent advancement of Large Language Models (LLMs) has established their potential as autonomous interactive agents. However, they often struggle in

agentsarxiv-cs-ai
14 Apr 2026
← Previous
1…138139140141142…375
Next →