AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,630Total entries
1Added by human
84,629Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,980 results
Model Releases

AutoMegaKernel: A Statically-Checked Agent Harness for Self-Retargeting Megakernel Synthesis

DGX agent

arXiv:2606.09682v1 Announce Type: new Abstract: AutoMegaKernel (AMK) compiles a HuggingFace Llama-family model into a single persistent cooperative CUDA kernel that runs the whole forward pass in one

model-releasesarxiv-cs-lg
9 Jun 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Safety

Baichuan-M4: A Clinical-Grade Medical Agent System for Continuous Care

DGX agent

arXiv:2606.08982v1 Announce Type: new Abstract: Baichuan-M4 is Baichuan Intelligence's clinical-grade medical large model, designed for continuous care rather than single-turn medical question answeri

safetyarxiv-cs-ai
9 Jun 2026
Safety

Cooperative Long Rope Skipping via Multi-Agent Reinforcement Learning

DGX agent

arXiv:2606.08064v1 Announce Type: new Abstract: Humans exhibit remarkable motor agility, enabling a wide range of dynamic skills such as running and jumping, which highlights the great potential of hu

safetyarxiv-cs-ro
9 Jun 2026
Safety

HARBOR: A Harness Framework for Agentic Robot Reinforcement Learning

DGX agent

arXiv:2606.08610v1 Announce Type: cross Abstract: Reinforcement learning (RL) has become a powerful paradigm for robot learning, particularly in sim-to-real settings, but its broader adoption remains

safetyarxiv-cs-ai
9 Jun 2026
Safety

HDRAgent: An Agentic Framework for Multi-Exposure HDR Imaging

DGX agent

arXiv:2606.09110v1 Announce Type: new Abstract: Most existing multi-exposure HDR methods follow a fixed feed-forward reconstruction paradigm, making them prone to ghosting artifacts in complex dynamic

safetyarxiv-cs-cv
9 Jun 2026
Model Releases

Principled Agent Debate: Adversarial Arbitration for Sycophancy Reduction in Large Language Models

DGX agent

arXiv:2606.07532v1 Announce Type: cross Abstract: RLHF-trained models are systematically biased toward agreement over accuracy, a structural property of the training process. We present Principled Age

model-releasesarxiv-cs-ai
9 Jun 2026
Agents

Syll: Open-Source Personal Automation with Cross-Surface Execution

DGX agent

arXiv:2606.07594v1 Announce Type: new Abstract: Personal AI agents must increasingly operate across APIs, shells, web surfaces, and desktop GUIs, yet many systems remain tuned to a single interface an

agentsarxiv-cs-ai
9 Jun 2026
Model Releases

MADE: Beyond Scoring via a Multilingual Agentic Diagnosing Engine for Fine-Grained Evaluation Insights

DGX agent

arXiv:2606.07020v1 Announce Type: new Abstract: Multilingual and multicultural benchmarks now cover dozens of languages and model families, but the resulting score landscapes remain metric-rich and in

model-releasesarxiv-cs-cl
8 Jun 2026
Model Releases

Agent-Orchestrated Adaptive RAG: A Comparative Study on Structured and Multi-Hop Retrieval

DGX agent

arXiv:2606.05658v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) enhances Large Language Models (LLMs) by grounding their responses in external knowledge, but conventional pipeli

model-releasesarxiv-cs-ai
6 Jun 2026
Hardware

CuTeGen: An LLM-Based Agentic Framework for Generation and Optimization of High-Performance GPU Kernels using CuTe

DGX agent

arXiv:2604.01489v2 Announce Type: replace-cross Abstract: High-performance GPU kernels are critical to modern machine learning systems, yet developing them remains a manual, expert-driven process. Rec

hardwarearxiv-cs-ai
6 Jun 2026
Model Releases

Vortex: Efficient and Programmable Sparse Attention Serving for AI Agents

DGX agent

arXiv:2606.06453v1 Announce Type: new Abstract: Sparse attention is becoming increasingly important for serving large language models (LLMs) as generation lengths continue to grow. However, deploying

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

Agents' Last Exam

DGX agent

arXiv:2606.05405v1 Announce Type: cross Abstract: Recent AI systems have achieved strong results on a wide range of benchmarks, yet these gains have not translated into economically meaningful deploym

model-releasesarxiv-cs-cl
5 Jun 2026
Model Releases

Thinking with Imagination: Agentic Visual Spatial Reasoning with World Simulators

DGX agent

arXiv:2606.06476v1 Announce Type: new Abstract: While Vision-Language Models (VLMs) have shown strong visual reasoning capabilities, their spatial reasoning abilities remain largely constrained to the

model-releasesarxiv-cs-cv
5 Jun 2026
Model Releases

CyberGym-E2E: Scalable Real-World Benchmark for AI Agents' End-to-End Cybersecurity Capabilities

DGX agent

arXiv:2606.04460v1 Announce Type: cross Abstract: AI has the potential to transform cybersecurity by enabling systems that can autonomously detect, analyze, and remediate software vulnerabilities. How

model-releasesarxiv-cs-ai
4 Jun 2026
Safety

SocialCoach: Personalized Social Skill Learning with RL-based Agentic Tutoring and Practice

DGX agent

arXiv:2606.04155v1 Announce Type: cross Abstract: Social skills such as negotiation and leadership are crucial for personal and professional success in today's interconnected world. However, scalable

safetyarxiv-cs-cl
4 Jun 2026
Model Releases

Solving Zebra Puzzles Using Constraint-Guided Multi-Agent Systems

DGX agent

arXiv:2407.03956v3 Announce Type: replace-cross Abstract: Prior research has enhanced the ability of Large Language Models (LLMs) to solve logic puzzles using techniques such as chain-of-thought promp

model-releasesarxiv-cs-cl
4 Jun 2026
Model Releases

What's new for Managed Service for Apache Spark clusters

DGX agent

At Google Cloud, our goal is to let you run large-scale analytical and data science workloads with maximum efficiency so you can process big data pipelines, machine learning, and ETL tasks. We recentl

model-releasesgoogle-cloud-ai
4 Jun 2026
Research

A Scoping Review of the Ethical Perspectives on Anthropomorphising Large Language Model-Based Conversational Agents

DGX agent

arXiv:2601.09869v2 Announce Type: replace Abstract: Anthropomorphisation -- the phenomenon whereby non-human entities are ascribed human-like qualities -- has become increasingly salient with the rise

researcharxiv-cs-ai
3 Jun 2026
Applications

Discovery to Execution: Scaling Agents with Toolboxes and Routines in Microsoft Foundry

DGX agent

Tooling doesn’t break at a small scale—it breaks when teams move to production. AI adoption accelerates, so does the number of tools available to them. Discovering, managing and securing the right too

applicationsmicrosoft-foundry
3 Jun 2026
Model Releases

From Prompt to Service: An SLM-Based Agent Orchestration Gateway for AI-Driven Virtual Worlds

DGX agent

arXiv:2606.03557v1 Announce Type: new Abstract: As generative AI capabilities expand, AI-driven virtual worlds face a growing architectural challenge. Users interact through in-world interfaces in mul

model-releasesarxiv-cs-ai
3 Jun 2026
Safety

Libra: Efficient Resource Management for Agentic RL Post-Training

DGX agent

arXiv:2606.03077v1 Announce Type: cross Abstract: Reinforcement learning (RL) has become a standard post-training paradigm for large language models (LLMs), extending beyond preference alignment to co

safetyarxiv-cs-ai
3 Jun 2026
Safety

Validation-Gated Multi-Agent Governance for Online Adaptation of Thermal-Hydraulic Surrogate Models under Operating-Regime Shift

DGX agent

arXiv:2606.03321v1 Announce Type: new Abstract: Artificial-intelligence surrogates can support second-by-second thermal-hydraulic forecasting, but models selected and frozen offline may become conditi

safetyarxiv-cs-lg
3 Jun 2026
Model Releases

VLESA: Vision-Language Embodied Safety Agent for Human Activity Monitoring

DGX agent

arXiv:2606.03954v1 Announce Type: new Abstract: As AI systems increasingly assist humans in physical tasks, ensuring safety becomes paramount -- physical actions carry immediate and irreversible conse

model-releasesarxiv-cs-cv
3 Jun 2026
Model Releases

3rd Place at CVPR 2026 CASTLE Challenge: Agentic Multi-View Long-Context Video Understanding via Hierarchical Knowledge Graph Retrieval

DGX agent

arXiv:2606.01933v1 Announce Type: new Abstract: This paper presents our winning methodology for the CASTLE 2026 Challenge at the CVPR 2026 EgoVis Workshop, where our team secured third place globally.

model-releasesarxiv-cs-cv
2 Jun 2026
Safety

Bridging the Last Mile of Time Series Forecasting with LLM Agents

DGX agent

arXiv:2606.02497v1 Announce Type: new Abstract: Time series forecasting has advanced rapidly, especially with the emergence of foundation models that show strong zero-shot performance on numerical ext

safetyarxiv-cs-ai
2 Jun 2026
Model Releases

CodeCytos: AI-assisted spatial molecular imaging analysis via code-augmented agent action space

DGX agent

arXiv:2606.00472v1 Announce Type: cross Abstract: Conventional tissue image analysis software provides foundational capabilities for cellular analysis, including segmentation, basic morphological feat

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

HALO: Learning Human-Robot Collaboration via Heterogeneous-Agent Lyapunov Policy Optimization

DGX agent

arXiv:2603.03741v2 Announce Type: replace-cross Abstract: To improve generalization and resilience in human-robot collaboration (HRC), robots must contend with diverse combinations of human behaviors

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

IMAC-AgriVLN: Can Agricultural Vision-and-Language Navigation Agents be Aware of Instruction Mistakes?

DGX agent

arXiv:2606.02519v1 Announce Type: new Abstract: Agricultural robots are serving as powerful assistants across a wide range of agricultural tasks, nevertheless, still heavily relying on manual operatio

model-releasesarxiv-cs-ro
2 Jun 2026
Model Releases

MCP-Persona: Benchmarking LLM Agents on Real-World Personal Applications via Environment Simulation

DGX agent

arXiv:2606.02470v1 Announce Type: new Abstract: The Model Context Protocol (MCP) has emerged as a transformative standard for connecting large language models (LLMs) with external data sources and too

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

RadioMaster: Multi-Agent System for Autonomous Radio Signal Generation

DGX agent

arXiv:2606.01862v1 Announce Type: cross Abstract: Translating user intents into physical radio signals represents the critical yet notoriously tedious final step in wireless prototyping, as it require

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Sandboxed Coding Agents are Competitive Omni-modal Task Solvers

DGX agent

arXiv:2606.00579v1 Announce Type: new Abstract: As multimodal LLMs increasingly target video and audio, it is often assumed that such tasks require native omnimodal models. We show that this is not al

model-releasesarxiv-cs-cl
2 Jun 2026
Safety

SceneSmith: Agentic Generation of Simulation-Ready Indoor Scenes

DGX agent

arXiv:2602.09153v2 Announce Type: replace-cross Abstract: Simulation has become a key tool for training and evaluating home robots at scale, yet existing environments fail to capture the diversity and

safetyarxiv-cs-ai
2 Jun 2026
Local Ai

An Organization-Scoped LLM Agent Runtime Architecture for Regulated Cybersecurity Operations

DGX agent

arXiv:2605.30604v1 Announce Type: cross Abstract: Regulated cybersecurity workflows lack a runtime substrate that enforces organization-level scope across retrieval, tool calls, memory, findings, repo

local-aiarxiv-cs-ai
1 Jun 2026
Model Releases

Crafter: A Multi-Agent Harness for Editable Scientific Figure Generation from Diverse Inputs

DGX agent

arXiv:2605.30611v1 Announce Type: cross Abstract: Scientific figures are among the most effective means of communicating complex research ideas, yet producing publication-quality illustrations remains

model-releasesarxiv-cs-ai
1 Jun 2026
Agents

More on the Hermes Skills Hub: https://hermes-agent.nousresearch.com/docs/guides/work-with-skills#the-skills-hub

DGX agent

The Hermes Skills Hub is a feature that allows users to discover, manage, and integrate skills within the Hermes agent framework, enabling extended functionality and customization of agent capabilitie

agentsnous-research--x
1 Jun 2026
Agents

Great article on harness engineering. https://www.langchain.com/blog/the-anatomy-of-an-agent-harness

DGX agent

This article from LangChain explores the architectural components and design principles of agent harnesses, which are systems that manage the execution and behavior of AI agents. The piece likely cove

agentsharrison-chase--x
31 May 2026
Industry

Personal agents light the fuse as Snowflake and Databricks move up the AI stack

DGX agent

The artificial intelligence wave is starting to look a bit like the personal computer era – with some obvious differences. The first similarity is personal productivity. Individuals are taking control

industrysiliconangle
30 May 2026
Agents

feel very aligned with our vision & Ronak + the awesome Trajectory team’s on practically tackling Continual Learning at scale 🚀 there’s a v…

DGX agent

feel very aligned with our vision & Ronak + the awesome Trajectory team’s on practically tackling Continual Learning at scale 🚀 there’s a very good reason why teams are partly building “Observability

agentsharrison-chase--x
29 May 2026
Safety

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation

DGX agent

arXiv:2605.29430v1 Announce Type: new Abstract: Automatic speech recognition (ASR) is a core component of human--computer interaction and an increasingly important front-end for LLM-based assistants a

safetyarxiv-cs-ai
29 May 2026
Safety

Train the Agent, Not the Expert: Learning to Harness Heterogeneous Experts for Multi-Turn Visual Reasoning

DGX agent

arXiv:2605.29894v1 Announce Type: new Abstract: Recent progress in computer vision has produced a wide range of powerful specialized models for detection, segmentation, counting, and other visual task

safetyarxiv-cs-cv
29 May 2026
Model Releases

Why Specialist Models Still Matter: A Heterogeneous Multi-Agent Paradigm for Medical Artificial Intelligence

DGX agent

arXiv:2605.29744v1 Announce Type: new Abstract: The impressive performance of generalist large language models (LLMs) such as GPT and Claude in healthcare raises a critical question: will domain-speci

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Chrome Enterprise rolls out AI agents and automation to streamline security management

DGX agent

Google LLC today launched new enhancements to Chrome Enterprise, the company’s enterprise version of its Chrome browser, designed to provide greater administrative and security control to information

model-releasessiliconangle
28 May 2026
Model Releases

MACReD: A Multi-Agent Collaborative Reasoning Framework for Reaction Diagram Parsing

DGX agent

arXiv:2605.28077v1 Announce Type: new Abstract: Parsing chemical reaction diagrams from scientific literature is challenging due to heterogeneous layouts, intertwined visual elements, and the difficul

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

MetaboT: An LLM-based Multi-Agent Frameworkfor Interactive Analysis of Mass SpectrometryMetabolomics Knowledge Graphs

DGX agent

arXiv:2510.01724v2 Announce Type: replace Abstract: Mass spectrometry-based metabolomics generates complex, high-dimensional data that holds vast potential for biological discovery but remains difficu

model-releasesarxiv-cs-ai
28 May 2026
Agents

Segment to Focus: Guiding Latent Action Models in the Presence of Distractors

DGX agent

arXiv:2602.02259v2 Announce Type: replace-cross Abstract: Latent action models (LAMs) offer a promising path to pre-training embodied agents on large amounts of action-free video. They infer latent ac

agentsarxiv-cs-cv
28 May 2026
Hardware

SwarmHarness: Skill-Based Task Routing via Decentralized Incentive-Aligned AI Agent Networks

DGX agent

arXiv:2605.28764v1 Announce Type: new Abstract: Vast quantities of compute (GPU cycles on personal workstations, idle inference servers, and edge devices between jobs) go unused because no incentive-a

hardwarearxiv-cs-ai
28 May 2026
Agents

Adversarial Training for Robust Coverage Network under Worst-case Facility Losses

DGX agent

arXiv:2605.26763v1 Announce Type: cross Abstract: The Maximal Covering Location-Interdiction Problem (MCLIP) is a classic bi-level optimization problem, which is fundamental to resilient infrastructur

agentsarxiv-cs-ai
27 May 2026
Model Releases

AutoDFT: A Closed-Loop Multi-Agent Framework for Autonomous DFT Calculations

DGX agent

arXiv:2605.26179v1 Announce Type: cross Abstract: Density functional theory (DFT) serves as the basis for computational discovery in materials science and chemistry, yet each calculation demands exten

model-releasesarxiv-cs-ai
27 May 2026
← Previous
1…174175176177178…375
Next →