AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “tools”

GridTimelineEvolution
5,141 results
Tutorials

Analyzing Multimodal Interaction Strategies for LLM-Assisted Manipulation of 3D Scenes

DGX agent

arXiv:2410.22177v2 Announce Type: replace-cross Abstract: As more applications of large language models (LLMs) for 3D content for immersive environments emerge, it is crucial to study user behaviour t

tutorialsarxiv-cs-ai
10 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

AV-SQL: Decomposing Complex Text-to-SQL Queries with Agentic Views

DGX agent

arXiv:2604.07041v1 Announce Type: cross Abstract: Text-to-SQL is the task of translating natural language queries into executable SQL for a given database, enabling non-expert users to access structur

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

Blind Refusal: Language Models Refuse to Help Users Evade Unjust, Absurd, and Illegitimate Rules

DGX agent

arXiv:2604.06233v1 Announce Type: new Abstract: Safety-trained language models routinely refuse requests for help circumventing rules. But not all rules deserve compliance. When users ask for help eva

model-releasesarxiv-cs-ai
10 Apr 2026
Agents

Can VLMs Unlock Semantic Anomaly Detection? A Framework for Structured Reasoning

DGX agent

arXiv:2510.18034v2 Announce Type: replace-cross Abstract: Autonomous driving systems remain critically vulnerable to the long-tail of rare, out-of-distribution semantic anomalies. While VLMs have emer

agentsarxiv-cs-ai
10 Apr 2026
Research

ConceptTracer: Interactive Analysis of Concept Saliency and Selectivity in Neural Representations

DGX agent

arXiv:2604.07019v1 Announce Type: cross Abstract: Neural networks deliver impressive predictive performance across a variety of tasks, but they are often opaque in their decision-making processes. Des

researcharxiv-cs-ai
10 Apr 2026
Safety

Data Leakage in Automotive Perception: Practitioners' Insights

DGX agent

arXiv:2604.06899v1 Announce Type: cross Abstract: Data leakage is the inadvertent transfer of information between training and evaluation datasets that poses a subtle, yet critical, risk to the reliab

safetyarxiv-cs-lg
10 Apr 2026
Safety

Decompose, Look, and Reason: Reinforced Latent Reasoning for VLMs

DGX agent

arXiv:2604.07518v1 Announce Type: new Abstract: Vision-Language Models often struggle with complex visual reasoning due to the visual information loss in textual CoT. Existing methods either add the c

safetyarxiv-cs-cl
10 Apr 2026
Model Releases

Distributed Interpretability and Control for Large Language Models

DGX agent

arXiv:2604.06483v1 Announce Type: cross Abstract: Large language models that require multiple GPU cards to host are usually the most capable models. It is necessary to understand and steer these model

model-releasesarxiv-cs-ai
10 Apr 2026
Tutorials

'Don't Be Afraid, Just Learn': Insights from Industry Practitioners to Prepare Software Engineers in the Age of Generative AI

DGX agent

arXiv:2604.06342v1 Announce Type: cross Abstract: Although tension between university curricula and industry expectations has existed in some form for decades, the rapid integration of generative AI (

tutorialsarxiv-cs-ai
10 Apr 2026
Model Releases

DosimeTron: Automating Personalized Monte Carlo Radiation Dosimetry in PET/CT with Agentic AI

DGX agent

arXiv:2604.06280v1 Announce Type: cross Abstract: Purpose: To develop and evaluate DosimeTron, an agentic AI system for automated patient-specific MC internal radiation dosimetry in PET/CT examination

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

E2Edev: Benchmarking Large Language Models in End-to-End Software Development Task

DGX agent

arXiv:2510.14509v3 Announce Type: replace-cross Abstract: The rapid advancement in large language models (LLMs) has demonstrated significant potential in End-to-End Software Development (E2ESD). Howev

model-releasesarxiv-cs-cl
10 Apr 2026
Research

Efficient PRM Training Data Synthesis via Formal Verification

DGX agent

arXiv:2505.15960v3 Announce Type: replace Abstract: Process Reward Models (PRMs) have emerged as a promising approach for improving LLM reasoning capabilities by providing process supervision over rea

researcharxiv-cs-cl
10 Apr 2026
Safety

EgoVerse: An Egocentric Human Dataset for Robot Learning from Around the World

DGX agent

arXiv:2604.07607v1 Announce Type: cross Abstract: Robot learning increasingly depends on large and diverse data, yet robot data collection remains expensive and difficult to scale. Egocentric human da

safetyarxiv-cs-cv
10 Apr 2026
Model Releases

Evaluating Repository-level Software Documentation via Question Answering and Feature-Driven Development

DGX agent

arXiv:2604.06793v1 Announce Type: cross Abstract: Software documentation is crucial for repository comprehension. While Large Language Models (LLMs) advance documentation generation from code snippets

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

Fail2Drive: Benchmarking Closed-Loop Driving Generalization

DGX agent

arXiv:2604.08535v1 Announce Type: cross Abstract: Generalization under distribution shift remains a central bottleneck for closed-loop autonomous driving. Although simulators like CARLA enable safe an

model-releasesarxiv-cs-cv
10 Apr 2026
Research

Faithful-First Reasoning, Planning, and Acting for Multimodal LLMs

DGX agent

arXiv:2511.08409v4 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) frequently suffer from unfaithfulness, generating reasoning chains that drift from visual evidence or contr

researcharxiv-cs-ai
10 Apr 2026
Model Releases

Fighting AI with AI: AI-Agent Augmented DNS Blocking of LLM Services during Student Evaluations

DGX agent

arXiv:2604.02360v1 Announce Type: cross Abstract: The transformative potential of large language models (LLMs) in education, such as improving accessibility and personalized learning, is being eclipse

model-releasesarxiv-cs-ai
10 Apr 2026
Research

Flemme: A Flexible and Modular Learning Platform for Medical Images

DGX agent

arXiv:2408.09369v3 Announce Type: replace-cross Abstract: As the rapid development of computer vision and the emergence of powerful network backbones and architectures, the application of deep learnin

researcharxiv-cs-cv
10 Apr 2026
Research

Floating or Suggesting Ideas? A Large-Scale Contrastive Analysis of Metaphorical and Literal Verb-Object Constructions

DGX agent

arXiv:2604.08275v1 Announce Type: new Abstract: Metaphor pervades everyday language, allowing speakers to express abstract concepts via concrete domains. While prior work has studied metaphors cogniti

researcharxiv-cs-cl
10 Apr 2026
Model Releases

From Fragments to Facts: A Curriculum-Driven DPO Approach for Generating Hindi News Veracity Explanations

DGX agent

arXiv:2507.05179v4 Announce Type: replace Abstract: In an era of rampant misinformation, generating reliable news explanations is vital, especially for under-represented languages like Hindi. Lacking

model-releasesarxiv-cs-cl
10 Apr 2026
Safety

Geometric Properties of the Voronoi Tessellation in Latent Semantic Manifolds of Large Language Models

DGX agent

arXiv:2604.06767v1 Announce Type: new Abstract: Language models operate on discrete tokens but compute in continuous vector spaces, inducing a Voronoi tessellation over the representation manifold. We

safetyarxiv-cs-lg
10 Apr 2026
Research

Improving Robustness In Sparse Autoencoders via Masked Regularization

DGX agent

arXiv:2604.06495v1 Announce Type: cross Abstract: Sparse autoencoders (SAEs) are widely used in mechanistic interpretability to project LLM activations onto sparse latent spaces. However, sparsity alo

researcharxiv-cs-ai
10 Apr 2026
Model Releases

KITE: Keyframe-Indexed Tokenized Evidence for VLM-Based Robot Failure Analysis

DGX agent

arXiv:2604.07034v1 Announce Type: cross Abstract: We present KITE, a training-free, keyframe-anchored, layout-grounded front-end that converts long robot-execution videos into compact, interpretable t

model-releasesarxiv-cs-ai
10 Apr 2026
Applications

Learning Geometry-Aware Nonprehensile Pushing and Pulling with Dexterous Hands

DGX agent

arXiv:2509.18455v4 Announce Type: replace Abstract: Nonprehensile manipulation, such as pushing and pulling, enables robots to move, align, or reposition objects that may be difficult to grasp due to

applicationsarxiv-cs-ro
10 Apr 2026
Research

Looking Beyond the Obvious: A Survey on Abstract Concept Recognition for Video Understanding

DGX agent

arXiv:2508.20765v2 Announce Type: replace-cross Abstract: The automatic understanding of video content is advancing rapidly. Empowered by deeper neural networks and large datasets, machines are increa

researcharxiv-cs-ai
10 Apr 2026
Agents

Mina: A Multilingual LLM-Powered Legal Assistant Agent for Bangladesh for Empowering Access to Justice

DGX agent

arXiv:2511.08605v3 Announce Type: replace Abstract: Bangladesh's low-income population faces major barriers to affordable legal advice due to complex legal language, procedural opacity, and high costs

agentsarxiv-cs-cl
10 Apr 2026
Research

Multi-modal user interface control detection using cross-attention

DGX agent

arXiv:2604.06934v1 Announce Type: cross Abstract: Detecting user interface (UI) controls from software screenshots is a critical task for automated testing, accessibility, and software analytics, yet

researcharxiv-cs-ai
10 Apr 2026
Research

MVOS_HSI: A Python Library for Preprocessing Agricultural Crop Hyperspectral Data

DGX agent

arXiv:2604.07656v1 Announce Type: cross Abstract: Hyperspectral imaging (HSI) allows researchers to study plant traits non-destructively. By capturing hundreds of narrow spectral bands per pixel, it r

researcharxiv-cs-cv
10 Apr 2026
Model Releases

Open-Ended Instruction Realization with LLM-Enabled Multi-Planner Scheduling in Autonomous Vehicles

DGX agent

arXiv:2604.08031v1 Announce Type: cross Abstract: Most Human-Machine Interaction (HMI) research overlooks the maneuvering needs of passengers in autonomous driving (AD). Natural language offers an int

model-releasesarxiv-cs-cv
10 Apr 2026
Safety

Qualixar OS: A Universal Operating System for AI Agent Orchestration

DGX agent

arXiv:2604.06392v1 Announce Type: new Abstract: We present Qualixar OS, the first application-layer operating system for universal AI agent orchestration. Unlike kernel-level approaches (AIOS) or sing

safetyarxiv-cs-ai
10 Apr 2026
Research

Reasoning Fails Where Step Flow Breaks

DGX agent

arXiv:2604.06695v1 Announce Type: new Abstract: Large reasoning models (LRMs) that generate long chains of thought now perform well on multi-step math, science, and coding tasks. However, their behavi

researcharxiv-cs-ai
10 Apr 2026
Safety

Reasoning Within the Mind: Dynamic Multimodal Interleaving in Latent Space

DGX agent

arXiv:2512.12623v3 Announce Type: replace-cross Abstract: Recent advancements in Multimodal Large Language Models (MLLMs) have significantly enhanced cross-modal understanding and reasoning by incorpo

safetyarxiv-cs-cl
10 Apr 2026
Model Releases

Riemann-Bench: A Benchmark for Moonshot Mathematics

DGX agent

arXiv:2604.06802v1 Announce Type: new Abstract: Recent AI systems have achieved gold-medal-level performance on the International Mathematical Olympiad, demonstrating remarkable proficiency at competi

model-releasesarxiv-cs-ai
10 Apr 2026
Safety

RoboAgent: Chaining Basic Capabilities for Embodied Task Planning

DGX agent

arXiv:2604.07774v1 Announce Type: cross Abstract: This paper focuses on embodied task planning, where an agent acquires visual observations from the environment and executes atomic actions to accompli

safetyarxiv-cs-cv
10 Apr 2026
Model Releases

SALLIE: Safeguarding Against Latent Language & Image Exploits

DGX agent

arXiv:2604.06247v1 Announce Type: cross Abstract: Large Language Models (LLMs) and Vision-Language Models (VLMs) remain highly vulnerable to textual and visual jailbreaks, as well as prompt injections

model-releasesarxiv-cs-ai
10 Apr 2026
Research

Sampling-Aware 3D Spatial Analysis in Multiplexed Imaging

DGX agent

arXiv:2604.07890v1 Announce Type: new Abstract: Highly multiplexed microscopy enables rich spatial characterization of tissues at single-cell resolution, yet most analyses rely on two-dimensional sect

researcharxiv-cs-cv
10 Apr 2026
Model Releases

sciwrite-lint: Verification Infrastructure for the Age of Science Vibe-Writing

DGX agent

arXiv:2604.08501v1 Announce Type: cross Abstract: Science currently offers two options for quality assurance, both inadequate. Journal gatekeeping claims to verify both integrity and contribution, but

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

SealQA: Raising the Bar for Reasoning in Search-Augmented Language Models

DGX agent

arXiv:2506.01062v4 Announce Type: replace Abstract: We introduce SealQA, a new challenge benchmark for evaluating SEarch-Augmented Language models on fact-seeking questions where web search yields con

model-releasesarxiv-cs-cl
10 Apr 2026
Agents

SkillClaw: Let Skills Evolve Collectively with Agentic Evolver

DGX agent

arXiv:2604.08377v1 Announce Type: cross Abstract: Large language model (LLM) agents such as OpenClaw rely on reusable skills to perform complex tasks, yet these skills remain largely static after depl

agentsarxiv-cs-cl
10 Apr 2026
Model Releases

Spectral Edge Dynamics Reveal Functional Modes of Learning

DGX agent

arXiv:2604.06256v1 Announce Type: cross Abstract: Training dynamics during grokking concentrate along a small number of dominant update directions -- the spectral edge -- which reliably distinguishes

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

The AI Skills Shift: Mapping Skill Obsolescence, Emergence, and Transition Pathways in the LLM Era

DGX agent

arXiv:2604.06906v1 Announce Type: cross Abstract: As Large Language Models reshape the global labor market, policymakers and workers need empirical data on which occupational skills may be most suscep

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

The Depth Ceiling: On the Limits of Large Language Models in Discovering Latent Planning

DGX agent

arXiv:2604.06427v1 Announce Type: cross Abstract: The viability of chain-of-thought (CoT) monitoring hinges on models being unable to reason effectively in their latent representations. Yet little is

model-releasesarxiv-cs-ai
10 Apr 2026
Research

The Unreasonable Effectiveness of Data for Recommender Systems

DGX agent

arXiv:2604.06420v2 Announce Type: cross Abstract: In recommender systems, collecting, storing, and processing large-scale interaction data is increasingly costly in terms of time, energy, and computat

researcharxiv-cs-lg
10 Apr 2026
Model Releases

Training Data Size Sensitivity in Unsupervised Rhyme Recognition

DGX agent

arXiv:2604.08156v1 Announce Type: new Abstract: Rhyme is deceptively intuitive: what is or is not a rhyme is constructed historically, scholars struggle with rhyme classification, and people disagree

model-releasesarxiv-cs-cl
10 Apr 2026
Research

Training-free Spatially Grounded Geometric Shape Encoding (Technical Report)

DGX agent

arXiv:2604.07522v1 Announce Type: new Abstract: Positional encoding has become the de facto standard for grounding deep neural networks on discrete point-wise positions, and it has achieved remarkable

researcharxiv-cs-cv
10 Apr 2026
Applications

Understanding Structured Financial Data with LLMs: A Case Study on Fraud Detection

DGX agent

arXiv:2512.13040v2 Announce Type: replace-cross Abstract: Detecting fraud in financial transactions typically relies on tabular models that demand heavy feature engineering to handle high-dimensional

applicationsarxiv-cs-cl
10 Apr 2026
Model Releases

Validated Intent Compilation for Constrained Routing in LEO Mega-Constellations

DGX agent

arXiv:2604.07264v1 Announce Type: cross Abstract: Operating LEO mega-constellations requires translating high-level operator intents ('reroute financial traffic away from polar links under 80 ms') int

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

VenusBench-Mobile: A Challenging and User-Centric Benchmark for Mobile GUI Agents with Capability Diagnostics

DGX agent

arXiv:2604.06182v1 Announce Type: cross Abstract: Existing online benchmarks for mobile GUI agents remain largely app-centric and task-homogeneous, failing to reflect the diversity and instability of

model-releasesarxiv-cs-ai
10 Apr 2026
← Previous
1…105106107108
Next →