AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,040 results
Model Releases

An Open-Source Benchmark and Baseline for Multi-temporal Referring Segmentation

DGX agent

arXiv:2606.00987v1 Announce Type: cross Abstract: Large Vision-Language Models (LVLMs) have shown strong visual understanding and language-guided grounding abilities, yet their capacity for multi-temp

model-releasesarxiv-cs-ai
2 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

ChartArena: Benchmarking Chart Parsing across Languages, Scenarios, and Formats

DGX agent

arXiv:2606.01348v1 Announce Type: new Abstract: Charts are a primary medium for conveying quantitative and relational information, yet systematically evaluating chart parsing models remains difficult.

model-releasesarxiv-cs-cv
2 Jun 2026
Safety

Civilizational Metamaterials: Engineering Coordination Under Capability Gradients and Structural Turbulence

DGX agent

arXiv:2606.00235v1 Announce Type: cross Abstract: We argue that governance must transition from a normative discipline to an engineering discipline, and we develop a formal framework, inspired by the

safetyarxiv-cs-ai
2 Jun 2026
Safety

Control of a Twin Rotor using Twin Delayed Deep Deterministic Policy Gradient (TD3)

DGX agent

arXiv:2512.13356v2 Announce Type: replace-cross Abstract: This paper proposes a reinforcement learning (RL) framework for controlling and stabilizing the Twin Rotor Aerodynamic System (TRAS) at specif

safetyarxiv-cs-ai
2 Jun 2026
Local Ai

DiscourseFlip: An Oblique Discourse-Level Opinion Manipulation Attack against Black-box Retrieval-Augmented Generation

DGX agent

arXiv:2606.01212v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) systems are widely deployed and increasingly influential, but their reliance on external corpora exposes new secu

local-aiarxiv-cs-ai
2 Jun 2026
Applications

Distributed GNEP Algorithms without Multiplier Sharing and Applications to Multi-Robot Coordination and Contextual Bandit-Based Active Learning

DGX agent

arXiv:2606.00759v1 Announce Type: new Abstract: Recent advances in artificial intelligence have expanded the focus from classical optimization to include equilibrium analysis in noncooperative games.

applicationsarxiv-cs-lg
2 Jun 2026
Safety

Expanding Spatial and Temporal Context for Robotic Imitation Learning With Scene Graphs

DGX agent

arXiv:2606.01072v1 Announce Type: cross Abstract: Imitation learning enables robots to learn how to execute tasks via observation. However, real-world environments like homes and offices are often sev

safetyarxiv-cs-cv
2 Jun 2026
Safety

From Cues to Horizons: Dynamic Risk Horizon Profiling for Trajectory Prediction

DGX agent

arXiv:2606.00857v1 Announce Type: cross Abstract: Accurate and reliable vehicle trajectory prediction is essential for safe autonomous driving. Recent studies have incorporated safety risk into trajec

safetyarxiv-cs-ai
2 Jun 2026
Model Releases

HAIM: Human-AI Music Datasets for AI Music Production Tracking Benchmark

DGX agent

arXiv:2606.01686v1 Announce Type: cross Abstract: As generative platforms such as Suno and Udio reach human-grade audio quality, the scope of AI's utility has expanded across the entire music producti

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

InPhyRe Discovers: Large Multimodal Models Struggle in Inductive Physical Reasoning

DGX agent

arXiv:2509.12263v3 Announce Type: replace Abstract: Large multimodal models (LMMs) encode physical laws observed during training, such as momentum conservation, as parametric knowledge. It allows LMMs

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

InstructSAM: Segment Any Instance with Any Instructions

DGX agent

arXiv:2605.26102v2 Announce Type: replace Abstract: In this paper, we introduce InstructSAM, a unified and streamlined framework designed for multi-instance segmentation under arbitrary instructions.

model-releasesarxiv-cs-cv
2 Jun 2026
Safety

Interpretable Policy Distillation for Power Grid Topology Control

DGX agent

arXiv:2606.00561v1 Announce Type: cross Abstract: Deep reinforcement learning (RL) offers a promising route to real-time power grid operation, yet large neural policies are costly to evaluate, hard to

safetyarxiv-cs-ai
2 Jun 2026
Model Releases

JenBridge: Adaptive Long-Form Video Soundtracking across Scene Transitions

DGX agent

arXiv:2606.01703v1 Announce Type: cross Abstract: We address the challenge of generating high-fidelity, long-form soundtracks that remain coherent across scene transitions. Existing AI music systems a

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Measuring and Mitigating Bias in Code Generated by Large Language Models

DGX agent

arXiv:2606.00049v1 Announce Type: cross Abstract: Large language models (LLMs) are widely recognised for their applications in natural language generation and are increasingly used for code generation

model-releasesarxiv-cs-ai
2 Jun 2026
Research

MOSS-Audio Technical Report

DGX agent

arXiv:2606.01802v1 Announce Type: cross Abstract: MOSS-Audio is a unified audio-language model for speech, environmental sound, and music understanding, supporting audio captioning, time-aware questio

researcharxiv-cs-ai
2 Jun 2026
Model Releases

OneVLA: A Unified Framework for Embodied Tasks

DGX agent

arXiv:2606.01241v1 Announce Type: new Abstract: Navigation and manipulation are fundamental capabilities of embodied intelligence, enabling robots to interpret natural language commands and interact p

model-releasesarxiv-cs-ro
2 Jun 2026
Model Releases

Personalized 3D Myocardial Infarct Geometry Reconstruction from Cine MRI for Cardiac Digital Twins

DGX agent

arXiv:2606.01808v1 Announce Type: new Abstract: Accurate 3D geometric characterization of myocardial infarction (MI) is essential for building cardiac digital twins (CDTs) to precisely simulate infarc

model-releasesarxiv-cs-cv
2 Jun 2026
Safety

Regime-Adaptive Continual Learning for Portfolio Management

DGX agent

arXiv:2606.00143v1 Announce Type: cross Abstract: Financial markets are inherently non-stationary, exhibiting frequent regime shifts and structural changes that render traditional Portfolio Management

safetyarxiv-cs-ai
2 Jun 2026
Model Releases

RoboStressBench: Benchmarking VLM Robustness to Physical Visual Stress in Embodied Scenes

DGX agent

arXiv:2606.00828v1 Announce Type: new Abstract: Vision-Language Models (VLMs) have shown strong visual understanding and are increasingly deployed in embodied AI systems, where reliable perception und

model-releasesarxiv-cs-cv
2 Jun 2026
Safety

Safety Alignment of LMs via Non-cooperative Games

DGX agent

arXiv:2512.20806v3 Announce Type: replace Abstract: Ensuring the safety of language models (LMs) while maintaining their usefulness remains a critical challenge in AI alignment. Current approaches rel

safetyarxiv-cs-ai
2 Jun 2026
Safety

Safety Game: Inference-Time Alignment of Black-Box LLMs via Constrained Optimization

DGX agent

arXiv:2510.09330v3 Announce Type: replace Abstract: Ensuring that large language models (LLMs) comply with safety requirements is a central challenge in AI deployment. Existing alignment approaches pr

safetyarxiv-cs-lg
2 Jun 2026
Safety

Scalable Ride-Sourcing Vehicle Rebalancing with Service Accessibility Guarantee: A Constrained Mean-Field Reinforcement Learning Approach

DGX agent

arXiv:2503.24183v3 Announce Type: replace Abstract: The expansion of ride-sourcing services such as Uber and Lyft has reshaped urban transportation by offering flexible, on-demand mobility via mobile

safetyarxiv-cs-lg
2 Jun 2026
Research

Shu Dao: A Calligraphy Score Framework Linking Calligraphy, Music, and Performance

DGX agent

arXiv:2606.00001v1 Announce Type: cross Abstract: This paper introduces Calligraphy Writing Score Representation (CWSR) and proposes Shu Dao as a framework that interprets East Asian calligraphy as a

researcharxiv-cs-cv
2 Jun 2026
Hardware

Skill-Based Mixture-of-Experts: Adaptive Routing for Heterogeneous Reasoning via Inferred Skills

DGX agent

arXiv:2503.05641v4 Announce Type: replace-cross Abstract: Combining existing pre-trained LLMs is a promising approach for diverse reasoning tasks. However, task-level expert selection is often too coa

hardwarearxiv-cs-ai
2 Jun 2026
Model Releases

SkyShield: Occupancy as a Safety Interface for Low-Altitude UAV Autonomy

DGX agent

arXiv:2606.00747v1 Announce Type: cross Abstract: For low-altitude Unmanned Aerial Vehicle (UAV) autonomy, 3D spatial understanding is not merely a perception objective, but the safety interface betwe

model-releasesarxiv-cs-ai
2 Jun 2026
Tutorials

Tackling the Root of Misinformation by Teaching Laypeople about Logical Fallacies via Socratic Questioning and Critical Argumentation

DGX agent

arXiv:2606.01020v1 Announce Type: new Abstract: Identifying logical fallacies in everyday discourse is challenging for many people. This challenge is amplified in the era of Large Language Models (LLM

tutorialsarxiv-cs-ai
2 Jun 2026
Safety

THRD: A Training-Free Multi-Turn Defense Framework for Jailbreak Attacks on Large Language Models

DGX agent

arXiv:2606.01738v1 Announce Type: cross Abstract: Multi-turn jailbreak attacks pose a growing threat to LLMs by exploiting conversational dynamics such as gradual escalation and cross-turn coordinatio

safetyarxiv-cs-ai
2 Jun 2026
Model Releases

Through the PRISM: Principle-Aware, Interpretable, and Multi-Scale Evaluation of Visual Designs

DGX agent

arXiv:2606.00592v1 Announce Type: new Abstract: Effective visual communication stems from the harmony of multiple design principles, such as readability, contrast, alignment, overlap, and coherence, w

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

Token Predictors Are Not Planners: Building Physically Grounded Causal Reasoners

DGX agent

arXiv:2606.01810v1 Announce Type: new Abstract: Current benchmarks for embodied vision-language planning often favor linguistic next-token prediction over physically grounded next-state reasoning. Thi

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Truthful AI Advisors: A Pre-Specified Benchmark for Large Language Model Honesty Under Preference Misalignment

DGX agent

arXiv:2606.01456v1 Announce Type: cross Abstract: Large language models are increasingly deployed as advisors whose objective is not aligned with the user's: recommenders optimize for engagement, sale

model-releasesarxiv-cs-cl
2 Jun 2026
Local Ai

VICR: Visual In-Context Restoration for Real-World Image Super-Resolution

DGX agent

arXiv:2606.00704v1 Announce Type: new Abstract: Real-world image super-resolution (Real-ISR) requires balancing structural fidelity to degraded observations with realistic detail synthesis. However, e

local-aiarxiv-cs-cv
2 Jun 2026
Model Releases

Where to Look: Can Foundation Models Reach a Target Viewpoint Through Active Exploration?

DGX agent

arXiv:2606.01247v1 Announce Type: new Abstract: Humans can reproduce the viewpoint specified by a target image through active head and body motion, yet spatial intelligence in foundation models has la

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

X-Stream: Exploring MLLMs as Multiplexers for Multi-Stream Understanding

DGX agent

arXiv:2606.02482v1 Announce Type: new Abstract: While video streaming understanding has made significant strides, real-world applications, such as live sports broadcasting, autonomous driving, and mul

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

Zero-Shot Off-Policy Learning

DGX agent

arXiv:2602.01962v2 Announce Type: replace-cross Abstract: Off-policy learning methods seek to derive an optimal policy directly from a fixed dataset of prior interactions. This objective presents sign

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

A Visually Impaired Assistance Benchmark for VLM-as-a-Judge Evaluation

DGX agent

arXiv:2605.31351v1 Announce Type: new Abstract: AI-based Visually Impaired Assistance (VIA) remains challenging, largely due to the high cost of human evaluation. The VLM-as-a-Judge paradigm may offer

model-releasesarxiv-cs-cl
1 Jun 2026
Safety

Can Aerial VLA Models Cooperate? Evaluating Closed-Loop Air-Ground Coordination with CARLA-Air

DGX agent

arXiv:2605.31066v1 Announce Type: new Abstract: Recent aerial vision-language-action (VLA) models show promising single-UAV capabilities, such as tracking moving objects and navigating to language-spe

safetyarxiv-cs-ro
1 Jun 2026
Model Releases

Can LLM Teams Play What? Where? When?

DGX agent

arXiv:2605.30459v1 Announce Type: new Abstract: Large language models (LLMs) remain limited on tasks requiring indirect reasoning, cultural knowledge, and coordinated hypothesis testing. We investigat

model-releasesarxiv-cs-cl
1 Jun 2026
Model Releases

Chain-of-Thought Reasoning In The Wild Is Not Always Faithful

DGX agent

arXiv:2503.08679v5 Announce Type: replace Abstract: Recent studies indicate that when faced with explicit biases in prompts, models often omit mentioning these biases in their Chain-of-Thought (CoT) o

model-releasesarxiv-cs-ai
1 Jun 2026
Research

Discovering Differences in Strategic Behavior Between Humans and LLMs

DGX agent

arXiv:2602.10324v2 Announce Type: replace Abstract: As Large Language Models (LLMs) are increasingly deployed in social and strategic scenarios, it becomes critical to understand where and why their b

researcharxiv-cs-ai
1 Jun 2026
Model Releases

DTBench: A Synthetic Benchmark for Document-to-Table Extraction

DGX agent

arXiv:2602.13812v3 Announce Type: replace-cross Abstract: Document-to-table (Doc2Table) extraction derives structured tables from unstructured documents under a target schema, enabling reliable and ve

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

GEM-Bench: A Benchmark for Ad-Injected Response Generation within Generative Engine Marketing

DGX agent

arXiv:2509.14221v3 Announce Type: replace-cross Abstract: Generative Engine Marketing (GEM) is an emerging ecosystem for monetizing generative engines, such as LLM-based chatbots, by seamlessly integr

model-releasesarxiv-cs-cl
1 Jun 2026
Safety

Healthcare Mechanisms from Policy-as-Code Search under Strategic Provider Response

DGX agent

arXiv:2605.30680v1 Announce Type: new Abstract: Healthcare mechanisms are inseparable from the strategic provider response they induce: existing healthcare AI benchmarks hold this response fixed and s

safetyarxiv-cs-ai
1 Jun 2026
Model Releases

MedFact: Benchmarking the Fact-Checking Capabilities of Large Language Models on Chinese Medical Texts

DGX agent

arXiv:2509.12440v3 Announce Type: replace-cross Abstract: Deploying Large Language Models (LLMs) in medical applications requires fact-checking capabilities to ensure patient safety and regulatory com

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

Mellum2 Technical Report

DGX agent

arXiv:2605.31268v1 Announce Type: new Abstract: We present Mellum 2, an open-weight 12B-parameter Mixture-of-Experts (MoE) language model with 2.5B active parameters per token. Mellum 2 is a general-p

model-releasesarxiv-cs-cl
1 Jun 2026
Applications

Neither Replacement nor Panacea: Comparing LLM-Based Conversational and Graphical Decision Support in Industrial Tasks

DGX agent

arXiv:2605.31287v1 Announce Type: cross Abstract: Managers in manufacturing settings rely on digital interfaces to interpret operational data for decision-making, but growing data volume and complexit

applicationsarxiv-cs-ai
1 Jun 2026
Model Releases

nuReasoning: A Reasoning-Centric Dataset and Benchmark for Long-Tail Autonomous Driving

DGX agent

arXiv:2605.31572v1 Announce Type: new Abstract: Reasoning is essential for autonomous driving (AD) in long-tail scenarios, where vehicles must apply commonsense knowledge, understand spatial relations

model-releasesarxiv-cs-cv
1 Jun 2026
Model Releases

SAW-Bench: Learning Situated Awareness in the Real World

DGX agent

arXiv:2602.16682v2 Announce Type: replace Abstract: A core aspect of human perception is situated awareness, the ability to relate ourselves to the surrounding physical environment and reason over pos

model-releasesarxiv-cs-cv
1 Jun 2026
Model Releases

Send a SCOUT First: Pre-hoc Reasoning for Adaptive Detector Allocation in Prompt-Injection Defense

DGX agent

arXiv:2605.30837v1 Announce Type: cross Abstract: Prompt-injection detectors are heterogeneous: each is strong on a different slice of attacks, and none is always reliable. Yet existing systems still

model-releasesarxiv-cs-lg
1 Jun 2026
← Previous
1…212213214215216…230
Next →