AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “research”

GridTimelineEvolution
25,629 results
26 May 2026

From Model Scaling to System Scaling: Scaling the Harness in Agentic AI

Model ReleasesDGX agent

arXiv:2605.26112v1 Announce Type: new Abstract: This paper studies the next major bottleneck in agentic AI as system scaling, not only model scaling: the design of auditable, persistent, modular, and

From Sleep Staging to Spindle Detection: A Case Study on End-to-End Automated Sleep Analysis

ApplicationsDGX agent

arXiv:2505.05371v2 Announce Type: replace-cross Abstract: Automation of sleep analysis, including both macrostructural (sleep stages) and microstructural (e.g., sleep spindles) elements, promises to e

Grammatically-Guided Sparse Attention for Efficient and Interpretable Transformers

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.24518v1 Announce Type: cross Abstract: The quadratic complexity of self-attention in Transformer models remains a significant bottleneck for processing long sequences and deploying large la

GroupTravelBench: Benchmarking LLM Agents on Multi-Person Travel Planning

Model ReleasesDGX agent

arXiv:2605.25200v1 Announce Type: new Abstract: Travel planning is a realistic task for evaluating the planning and tool-use abilities of LLM agents. However, existing benchmarks typically assume only

Grow-Prune-Freeze Networks: Adaptive & Continual Learning Technique for Olfactory Navigation

SafetyDGX agent

arXiv:2605.25170v1 Announce Type: cross Abstract: Training data for olfaction is scattered through disparate, non-standardized datasets that limit the ability to build representative world models. Olf

Hide-and-Shill: A Reinforcement Learning Framework for Market Manipulation Detection in Symphony-a Decentralized Multi-Agent System

SafetyDGX agent

arXiv:2507.09179v3 Announce Type: replace Abstract: Decentralized finance (DeFi) has introduced a new era of permissionless financial innovation but also led to unprecedented market manipulation. With

Improving the Completeness and Comparability of Segment Disclosures: A Large Language Model Approach

SafetyDGX agent

arXiv:2605.23924v1 Announce Type: new Abstract: Segment-level disclosures are a central component of financial reporting, providing insight into firms' internal organization and the allocation of econ

JT-SAFE-V2: Safety-by-Design Foundation Model with World-Context Data

SafetyDGX agent

arXiv:2605.24414v1 Announce Type: new Abstract: We introduce JT-Safe-V2, a large language model designed to advance the safety and trustworthiness of foundation models, extending our previous JT-Safe

JudgmentBench: Comparing Rubric and Preference Evaluation for Quality Assessment

Model ReleasesDGX agent

arXiv:2605.25240v1 Announce Type: cross Abstract: Two methodologies dominate current practices of benchmarking: rubric-based scoring evaluates items against predefined criteria, whereas comparative ju

LLM Agent Based Renewable Energy Forecasting Using Edge and IoT Data A Review of Solar Wind Weather and Grid Aware Decision Support

Local AiDGX agent

arXiv:2605.25141v1 Announce Type: cross Abstract: Reliable forecasting of renewable energy generation is a foundational requirement for grid stability energy trading battery scheduling and carbon awar

MCPXKIT: The Unified Toolkit for Analyzing Model Context Protocol Security

AgentsDGX agent

arXiv:2508.12538v2 Announce Type: replace-cross Abstract: The Model Context Protocol (MCP) has emerged as a universal standard that enables AI agents to seamlessly connect with external tools, signifi

Meta-Soft: Leveraging Composable Meta-Tokens for Context-Preserving KV Cache Compression

Model ReleasesDGX agent

arXiv:2605.22337v2 Announce Type: replace Abstract: The KV cache used in large language models has linearly growing time complexity, so LLMs face memory blow-up and reduced decoding efficiency when th

MIND: Multi-Scale Intent Diffusion for Text-Driven Physics-Based Humanoid Control

Model ReleasesDGX agent

arXiv:2605.26006v1 Announce Type: cross Abstract: Enabling physics-based humanoids to execute diverse behaviors from high-level textual commands remains a significant challenge. Existing methods typic

Multi-market value-stacking: Battery control for combined imbalance participation and non-uniform FCR bidding

AgentsDGX agent

arXiv:2605.23964v1 Announce Type: cross Abstract: The growing share of Renewable Energy Sources (RES) in modern power systems increases both grid imbalances and frequency deviations, reinforcing the n

Multi-Persona Debate System for Automated Scientific Hypothesis Generation

AgentsDGX agent

arXiv:2605.23917v1 Announce Type: new Abstract: Modern scientific discovery is bottlenecked not by data scarcity, but by the inability to synthesize fragmented knowledge into actionable hypotheses. Th

MuNet: A Mutualistic Network for Joint 3D Human Mesh Recovery and 3D Clothed Human Reconstruction from Single Images

Model ReleasesDGX agent

arXiv:2605.25861v1 Announce Type: cross Abstract: 3D human mesh recovery and 3D clothed human reconstruction are inherently related, yet they have long been studied in isolation, thereby overlooking t

New study: Securing AI in the browser is a top priority for IT Leaders

TutorialsDGX agent

The way we work has fundamentally changed. From automated agents to sophisticated AI services, Generative AI (GenAI) has become a daily tool for a vast majority of employees. But with this rapid adopt

Non-Invasive Reconstruction of Intracranial EEG Across the Deep Temporal Lobe from Scalp EEG based on Conditional Normalizing Flow

Local AiDGX agent

arXiv:2603.03354v3 Announce Type: replace-cross Abstract: Although obtaining deep brain activity from non-invasive scalp electroencephalography (sEEG) is crucial for neuroscience and clinical diagnosi

Parameter-Efficient VLMs for Gastrointestinal Endoscopy: Medical Image Generation and Clinical Visual Question Answering

Model ReleasesDGX agent

arXiv:2605.24792v1 Announce Type: cross Abstract: The major limitations of gastrointestinal (GI) endoscopy AI systems arise from a shortage of annotated data, strict privacy policies, and significant

Positivity in classical enumerative geometry: a case study in synchronized AI-assisted mathematics

ApplicationsDGX agent

arXiv:2605.25271v1 Announce Type: cross Abstract: We study the symmetric polynomial prod_{alphain A_{n,d}}igl(1+alpha_1 x_1+dots+alpha_n x_nigr) where A_{n,d}:={alphainZ_{ge 0}^n:|alpha|=d}, which is

Practical Quantum CIM Empowerment via All-Domestic-Core Agentic Large Model

AgentsDGX agent

arXiv:2605.23934v1 Announce Type: new Abstract: Quantum computing devices are recognized as powerful tools for solving NP-complete problems. However, the intricacy of their modeling presents notable b

RAW: Robust Avatar Watermarking -- Benchmarking and Baseline

Model ReleasesDGX agent

arXiv:2605.23994v1 Announce Type: cross Abstract: Digital avatar watermarking presents unique challenges: avatars are routinely post-processed with background replacement, reframing, and format conver

Scaling up Energy-Aware Multi-Agent Reinforcement Learning for Mission-Oriented Drone Networks with Individual Reward

AgentsDGX agent

arXiv:2605.24992v1 Announce Type: cross Abstract: Multi-agent reinforcement learning (MARL) has shown wide applicability in collaborative systems such as autonomous driving and smart cities for its ab

Security in the Fine-Tuning Lifecycle of Large Language Models: Threats, Defenses,Evaluation, and Future Directions

Model ReleasesDGX agent

arXiv:2605.25073v1 Announce Type: cross Abstract: Background: Fine-tuning is central to adapting pre-trained Large Language Models (LLMs) to downstream tasks, but its reliance on training data, parame

Side-by-side Comparison Amplifies Dialect Bias in Language Models

SafetyDGX agent

arXiv:2605.24384v1 Announce Type: cross Abstract: Language models (LMs) can exhibit systematic biases against speakers based on variations in their dialects, even in the absence of a dialect label, a

Software Engineer Roles @SakanaAILabs https://sakana.ai/careers/#software-engineer-enterprise

ApplicationsDGX agent

Sakana AI Labs is recruiting software engineers for enterprise-focused roles, with details available on their careers page. The position likely involves developing and maintaining software infrastruct

SoK: DARPA's AI Cyber Challenge (AIxCC): Competition Design, Architectures, and Lessons Learned

AgentsDGX agent

arXiv:2602.07666v3 Announce Type: replace-cross Abstract: DARPA's AI Cyber Challenge (AIxCC, 2023--2025) is the largest competition to date for building fully autonomous cyber reasoning systems (CRSs)

System scaling is the next real bottleneck in agentic AI. If you build agent orchestration layers, this is a clean map of where the engineer…

Model ReleasesDGX agent

System scaling is the next real bottleneck in agentic AI. If you build agent orchestration layers, this is a clean map of where the engineering leverage actually sits. The labs own the model. You own

Teaching large language models to reason like expert diagnosticians

Model ReleasesDGX agent

arXiv:2509.12194v2 Announce Type: replace Abstract: Differential diagnosis is an iterative process that integrates patient information with broader medical knowledge. Clinical case series such as the

The Perception-Physics Paradox: Probing Scientific Alignment with TC-Bench

Model ReleasesDGX agent

arXiv:2605.24782v1 Announce Type: new Abstract: While Vision Foundation Models (VFMs) excel at predictive tasks on satellite imagery, their performance can arise from visual correlations rather than u

The pressure

ToolsDGX agent

The pressure Daniel Stenberg on the unprecedented level of pressure the curl team are facing right now thanks to the deluge of (credible) AI-assisted security issues being reported. The rate of incomi

Toward a Benchmark for Controllable Simulation of Imperfect Students with Large Language Models

Model ReleasesDGX agent

arXiv:2605.25601v1 Announce Type: cross Abstract: Teacher education requires deliberate practice with learners who exhibit identifiable strengths, weaknesses, and partial mastery. Large language model

Towards Cognitively-Faithful Decision-Making Models to Improve AI Alignment

SafetyDGX agent

arXiv:2509.04445v2 Announce Type: replace Abstract: Recent AI trends seek to align AI models to learned human-centric objectives, such as personal preferences, utility, or societal values. Using stand

Towards trustworthy agentic AI: a comprehensive survey of safety, robustness, privacy, and system security

SafetyDGX agent

arXiv:2605.23989v1 Announce Type: new Abstract: Agentic AI systems -- Large Language Models (LLMs) augmented with planning, tool use, memory, and long-horizon interactions -- can execute complex tasks

TRACE: A taxonomy-grounded synthetic dataset for teaching-program generation and session interpretation in Applied Behavior Analysis

Model ReleasesDGX agent

arXiv:2605.25038v1 Announce Type: new Abstract: Applied Behavior Analysis (ABA) is a clinical discipline whose documentation, teaching programs and multi-session behavioral logs, is formulaic and high

TriVAL: A Tri-Validation Framework for Faithful Automatic Optimization Modeling

Model ReleasesDGX agent

arXiv:2605.23966v1 Announce Type: cross Abstract: Optimization modeling serves as the pivotal bridge between natural-language problem descriptions and optimization solvers, and remains a cornerstone f

Understanding Conversational Patterns in Multi-agent Programming: A Case Study on Fibonacci Game Development

Model ReleasesDGX agent

arXiv:2605.24138v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly applied to software engineering (SE), yet their potential for autonomous, role-oriented collaboration re

Unlocking Apple's Private Cloud Compute: An Analysis of Privacy-Preserving Artificial Intelligence

Local AiDGX agent

arXiv:2605.24239v1 Announce Type: cross Abstract: Many existing Artificial Intelligence (AI) solutions on mobile devices rely on an extensive collection of sensitive data, raising privacy concerns and

ViroBench: Benchmarking Nucleotide Foundation Models on Viral Genomics Tasks

Model ReleasesDGX agent

arXiv:2605.25388v1 Announce Type: new Abstract: Nucleotide sequences constitute the fundamental genetic basis of biological systems, rendering viral genomic analysis critical for biomedical advancemen

Who judges the judges? Governance from metrics: a runtime framework for continuous LLM compliance monitoring

Model ReleasesDGX agent

arXiv:2605.24737v1 Announce Type: cross Abstract: Current approaches to AI compliance treat conformity as a binary, audit-time verdict rather than a continuous, measurable property of production syste

WideDepth: Millimeter-Accurate Benchmark for Fisheye Depth Estimation

Model ReleasesDGX agent

arXiv:2605.24074v1 Announce Type: cross Abstract: Fisheye cameras are increasingly adopted in robotics for near-field manipulation, navigation, and immersive perception, yet indoor depth benchmarks wi

World-State Transformations for Neuro-symbolic Interactive Storytelling

Model ReleasesDGX agent

arXiv:2605.24719v1 Announce Type: cross Abstract: Large Language Models (LLMs) have changed the possibilities of Interactive Storytelling systems that process free-text user input. However, as more of

25 May 2026

6G Communication Networks Enabling Embodied Agents: Architecture and Prototype

AgentsDGX agent

arXiv:2605.23263v1 Announce Type: cross Abstract: Embodied agents, which couple intelligent decision-making with physical actuation in the real world, impose far more stringent and heterogeneous commu

A Comparative Evaluation of Structural Topic Models and BERTopic for Short, Open-Ended Survey Responses

Model ReleasesDGX agent

arXiv:2605.23093v1 Announce Type: new Abstract: Topic modeling in applied psychology increasingly spans two methodological traditions: probabilistic bag-of-words models and newer embedding-based appro

A full tour through RAG, document context, and AI agents - from 2023 to 2026 🌎🤖 @hexapode gave a comprehensive 90-min workshop at @aiDotEn…

AgentsDGX agent

A full tour through RAG, document context, and AI agents - from 2023 to 2026 🌎🤖 @hexapode gave a comprehensive 90-min workshop at @aiDotEngineer Singapore last week that comprehensively traces through

A Novel Approach for the Counting of Wood Logs Using cGANs and Image Processing Techniques

SafetyDGX agent

arXiv:2605.23775v1 Announce Type: new Abstract: This study tackles the challenge of precise wood log counting, where applications of the proposed methodology can span from automated approaches for mat

AI Evaluation Should Require Standardized Item-Level Data Releases

Model ReleasesDGX agent

arXiv:2604.03244v2 Announce Type: replace Abstract: This position paper argues that standardized item-level benchmark data should become the default infrastructure for AI evaluation. Current evaluatio

An Open-Source Training Dataset for Foundation Models for Black-box Optimization

TutorialsDGX agent

arXiv:2605.23417v1 Announce Type: new Abstract: Most black-box optimization methods require extensive hyperparameter tuning, often limiting their ability to generalize across different optimization do

Co-ReAct: Rubrics as Step-Level Collaborators for ReAct Agents

AgentsDGX agent

arXiv:2605.23590v1 Announce Type: new Abstract: ReAct-style agents for search-intensive, multi-step reasoning tasks rely largely on their own internal judgment to decide what evidence to seek, which r

CultivAgents: Cultivating Relationship-Centered Multi-Agent Systems for Personalized Gardening

Local AiDGX agent

arXiv:2605.23193v1 Announce Type: cross Abstract: Gardening is critical to support well-being, cultural continuity, and food autonomy, yet existing digital tools often provide generic advice that over

Curriculum reinforcement learning with measurable task representation learning

AgentsDGX agent

arXiv:2605.23372v1 Announce Type: cross Abstract: In curriculum reinforcement learning (CRL), an agent incrementally accumulates knowledge over a sequence of tasks (i.e., a curriculum), and the learni

DRIVESPATIAL: A Benchmark for Spatiotemporal Intelligence in VLMs for Autonomous Driving

Model ReleasesDGX agent

arXiv:2605.23176v1 Announce Type: new Abstract: Spatiotemporal intelligence in autonomous driving (AD) requires an agent to integrate multi-view observations into a coherent scene representation, main

Emotion Recognition in Sign Language Conversation

ApplicationsDGX agent

arXiv:2605.23328v1 Announce Type: new Abstract: Emotion Recognition in Conversation is a core component of affective computing, while current resources of sign language emotion datasets primarily focu

Empowering 9-1-1 Calltaking Training with Generative AI: Experiences and Lessons Learned

SafetyDGX agent

arXiv:2602.13241v2 Announce Type: replace-cross Abstract: Emergency call-takers form the first operational link in public safety response, handling over 240 million calls annually while facing a susta

Evaluating Large Language Models in a Complex Hidden Role Game

Model ReleasesDGX agent

arXiv:2605.22826v1 Announce Type: cross Abstract: Quantifying the deceptive potential of Large Language Models (LLMs) is critical for AI safety, yet difficult to achieve in uncontrolled environments.

Evaluating Memory Structure in LLM Agents

Model ReleasesDGX agent

arXiv:2602.11243v2 Announce Type: replace-cross Abstract: Modern LLM-based agents and chat assistants rely on long-term memory frameworks to store reusable knowledge, recall user preferences, and augm

Exploring deep learning for Event-Based Saliency Prediction with a Transformer-based model

Model ReleasesDGX agent

arXiv:2605.23790v1 Announce Type: new Abstract: Saliency prediction has been extensively studied in RGB images and videos as a computational model of human visual attention. In contrast, predicting sa

Four Simple Proprioceptive Estimators for Legged Robots

SafetyDGX agent

arXiv:2605.23100v1 Announce Type: new Abstract: Legged robots carry an IMU, but the inertial solution drifts because consumer-grade IMUs are noisy. However, the feet create intermittent contacts with

GeoMAE: Masking Representation Learning for Spatio-Temporal Graph Forecasting with Missing Values

ApplicationsDGX agent

arXiv:2508.14083v3 Announce Type: replace-cross Abstract: The ubiquity of missing data in urban intelligence systems, attributable to adverse environmental conditions and equipment failures, poses a s

Harness, Scaffold, and the AI Agent Terms Worth Getting Right

AgentsDGX agent

This article defines and clarifies key terminology related to AI agents, including the concepts of 'harness' and 'scaffold,' which are important architectural and operational components in building an

← Previous
1…399400401402403…428
Next →