AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “tools”

GridTimelineEvolution
5,141 results
Model Releases

VLA Foundry: A Unified Framework for Training Vision-Language-Action Models

DGX agent

arXiv:2604.19728v1 Announce Type: cross Abstract: We present VLA Foundry, an open-source framework that unifies LLM, VLM, and VLA training in a single codebase. Most open-source VLA efforts specialize

model-releasesarxiv-cs-ai
22 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents

A Discordance-Aware Multimodal Framework with Multi-Agent Clinical Reasoning

DGX agent

arXiv:2604.16333v1 Announce Type: new Abstract: Knee osteoarthritis frequently exhibits discordance between structural damage observed in imaging and patient-reported symptoms such as pain. This misma

agentsarxiv-cs-lg
21 Apr 2026
Research

A Two-Stage Deep Learning Framework for Segmentation of Ten Gastrointestinal Organs from Coronal MR Enterography

DGX agent

arXiv:2604.17118v1 Announce Type: cross Abstract: Accurate segmentation of gastrointestinal (GI) organs in magnetic resonance enterography (MRE) is critical for diagnosing inflammatory bowel disease (

researcharxiv-cs-cv
21 Apr 2026
Model Releases

Agentic Risk-Aware Set-Based Engineering Design

DGX agent

arXiv:2604.16687v1 Announce Type: cross Abstract: This paper introduces a multi-agent framework guided by Large Language Models (LLMs) to assist in the early stages of engineering design, a phase ofte

model-releasesarxiv-cs-lg
21 Apr 2026
Agents

Agents Explore but Agents Ignore: LLMs Lack Environmental Curiosity

DGX agent

arXiv:2604.17609v1 Announce Type: new Abstract: LLM-based agents are assumed to integrate environmental observations into their reasoning: discovering highly relevant but unexpected information should

agentsarxiv-cs-cl
21 Apr 2026
Safety

Alignment Data Map for Efficient Preference Data Selection and Diagnosis

DGX agent

arXiv:2505.23114v3 Announce Type: replace Abstract: Human preference data is essential for aligning large language models (LLMs) with human values, but collecting such data is often costly and ineffic

safetyarxiv-cs-cl
21 Apr 2026
Model Releases

BenchMarker: An Education-Inspired Toolkit for Highlighting Flaws in Multiple-Choice Benchmarks

DGX agent

arXiv:2602.06221v2 Announce Type: replace Abstract: Multiple-choice question answering (MCQA) is standard in NLP, but benchmarks lack rigorous quality control. We present BenchMarker, an education-ins

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Beyond 'I Don't Know': Evaluating LLM Self-Awareness in Discriminating Data and Model Uncertainty

DGX agent

arXiv:2604.17293v1 Announce Type: new Abstract: Reliable Large Language Models (LLMs) should abstain when confidence is insufficient. However, prior studies often treat refusal as a generic 'I don't k

model-releasesarxiv-cs-cl
21 Apr 2026
Research

Beyond Static Benchmarks: Synthesizing Harmful Content via Persona-based Simulation for Robust Evaluation

DGX agent

arXiv:2604.17020v1 Announce Type: new Abstract: Static benchmarks for harmful content detection face limitations in scalability and diversity, and may also be affected by contamination from web-scale

researcharxiv-cs-cl
21 Apr 2026
Research

BhashaSutra: A Task-Centric Unified Survey of Indian NLP Datasets, Corpora, and Resources

DGX agent

arXiv:2604.18423v1 Announce Type: new Abstract: India's linguistic landscape, spanning 22 scheduled languages and hundreds of marginalized dialects, has driven rapid growth in NLP datasets, benchmarks

researcharxiv-cs-cl
21 Apr 2026
Safety

BIASEDTALES-ML: A Multilingual Dataset for Analyzing Narrative Attribute Distributions in LLM-Generated Stories

DGX agent

arXiv:2604.17008v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly used to generate narrative content, including children's stories, which play an important role in social a

safetyarxiv-cs-cl
21 Apr 2026
Local Ai

Bridging the Culture Gap: A Framework for LLM-Driven Socio-Cultural Localization of Math Word Problems in Low-Resource Languages

DGX agent

arXiv:2508.14913v4 Announce Type: replace Abstract: Large language models (LLMs) have demonstrated significant capabilities in solving mathematical problems expressed in natural language. However, mul

local-aiarxiv-cs-cl
21 Apr 2026
Applications

Can we generate portable representations for clinical time series data using LLMs?

DGX agent

arXiv:2603.23987v2 Announce Type: replace Abstract: Deploying clinical ML is slow and brittle: models that work at one hospital often degrade under distribution shifts at the next. In this work, we st

applicationsarxiv-cs-lg
21 Apr 2026
Research

Chaos-Enhanced Prototypical Networks for Few-Shot Medical Image Classification

DGX agent

arXiv:2604.17300v1 Announce Type: cross Abstract: The scarcity of labeled clinical data in oncology makes Few-Shot Learning (FSL) a critical framework for Computer Aided Diagnostics, but we observed t

researcharxiv-cs-cv
21 Apr 2026
Model Releases

ClawEnvKit: Automatic Environment Generation for Claw-Like Agents

DGX agent

arXiv:2604.18543v1 Announce Type: cross Abstract: Constructing environments for training and evaluating claw-like agents remains a manual, human-intensive process that does not scale. We argue that wh

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

CoDial: Interpretable Task-Oriented Dialogue Systems Through Dialogue Flow Alignment

DGX agent

arXiv:2506.02264v3 Announce Type: replace Abstract: Building Task-Oriented Dialogue (TOD) systems that generalize across different tasks remains a challenging problem. Data-driven approaches often str

model-releasesarxiv-cs-cl
21 Apr 2026
Safety

COSEARCH: Joint Training of Reasoning and Document Ranking via Reinforcement Learning for Agentic Search

DGX agent

arXiv:2604.17555v1 Announce Type: cross Abstract: Agentic search -- the task of training agents that iteratively reason, issue queries, and synthesize retrieved information to answer complex questions

safetyarxiv-cs-cl
21 Apr 2026
Model Releases

DEM Refinement and Validation on the Lunar Surface Using Shape-from-Shading with Chandrayaan-2 OHRC Imagery

DGX agent

arXiv:2604.17436v1 Announce Type: new Abstract: This study presents a Shape from Shading (SfS) framework to enhance sub-metre resolution lunar digital elevation models (DEMs) using imagery from the Or

model-releasesarxiv-cs-cv
21 Apr 2026
Safety

DreamShot: Personalized Storyboard Synthesis with Video Diffusion Prior

DGX agent

arXiv:2604.17195v1 Announce Type: new Abstract: Storyboard synthesis plays a crucial role in visual storytelling, aiming to generate coherent shot sequences that visually narrate cinematic events with

safetyarxiv-cs-cv
21 Apr 2026
Research

EmoVerse: A MLLMs-Driven Emotion Representation Dataset for Interpretable Visual Emotion Analysis

DGX agent

arXiv:2511.12554v2 Announce Type: replace Abstract: Visual Emotion Analysis (VEA) aims to bridge the affective gap between visual content and human emotional responses. Despite its promise, progress i

researcharxiv-cs-cv
21 Apr 2026
Research

Enabling Stroke-Level Structural Analysis of Hieroglyphic Scripts without Language-Specific Priors

DGX agent

arXiv:2601.05508v2 Announce Type: replace-cross Abstract: Hieroglyphs, as logographic writing systems, encode rich semantic and cultural information within their internal structural composition. Yet,

researcharxiv-cs-cl
21 Apr 2026
Safety

End-to-End Optimization of LLM-Driven Multi-Agent Search Systems via Heterogeneous-Group-Based Reinforcement Learning

DGX agent

arXiv:2506.02718v2 Announce Type: replace Abstract: Large language models (LLMs) are versatile, yet their deployment in complex real-world settings is limited by static knowledge cutoffs and the diffi

safetyarxiv-cs-lg
21 Apr 2026
Research

Enhancing Trust in Large Language Models via Uncertainty-Calibrated Fine-Tuning

DGX agent

arXiv:2412.02904v2 Announce Type: replace Abstract: Large language models (LLMs) have revolutionized the field of natural language processing with their impressive reasoning and question-answering cap

researcharxiv-cs-cl
21 Apr 2026
Model Releases

ESsEN: Training Compact Discriminative Vision-Language Transformers in a Low-Resource Setting

DGX agent

arXiv:2604.18452v1 Announce Type: cross Abstract: Vision-language modeling is rapidly increasing in popularity with an ever expanding list of available models. In most cases, these vision-language mod

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

FLiP: Towards understanding and interpreting multimodal multilingual sentence embeddings

DGX agent

arXiv:2604.18109v1 Announce Type: new Abstract: This paper presents factorized linear projection (FLiP) models for understanding pretrained sentence embedding spaces. We train FLiP models to recover t

model-releasesarxiv-cs-cl
21 Apr 2026
Research

From Fallback to Frontline: When Can LLMs be Superior Annotators of Human Perspectives?

DGX agent

arXiv:2604.17968v1 Announce Type: cross Abstract: Although large language models (LLMs) are increasingly used as annotators at scale, they are typically treated as a pragmatic fallback rather than a f

researcharxiv-cs-cl
21 Apr 2026
Research

FSEVAL: Feature Selection Evaluation Toolbox and Dashboard

DGX agent

arXiv:2604.18227v1 Announce Type: new Abstract: Feature selection is a fundamental machine learning and data mining task, involved with discriminating redundant features from informative ones. It is a

researcharxiv-cs-lg
21 Apr 2026
Tutorials

Graph neural network for colliding particles with an application to sea ice floe modeling

DGX agent

arXiv:2602.16213v2 Announce Type: replace-cross Abstract: This paper introduces a novel approach to sea ice modeling using Graph Neural Networks (GNNs), utilizing the natural graph structure of sea ic

tutorialsarxiv-cs-cv
21 Apr 2026
Research

HopWeaver: Cross-Document Synthesis of High-Quality and Authentic Multi-Hop Questions

DGX agent

arXiv:2505.15087v3 Announce Type: replace Abstract: Multi-Hop Question Answering (MHQA) is crucial for evaluating the model's capability to integrate information from diverse sources. However, creatin

researcharxiv-cs-cl
21 Apr 2026
Research

How Much Data is Enough? The Zeta Law of Discoverability in Biomedical Data, featuring the enigmatic Riemann zeta function

DGX agent

arXiv:2604.17581v1 Announce Type: new Abstract: How much data is enough to make a scientific discovery? As biomedical datasets scale to millions of samples and AI models grow in capacity, progress inc

researcharxiv-cs-lg
21 Apr 2026
Research

Hyperspectral Unmixing Hierarchies

DGX agent

arXiv:2604.16969v1 Announce Type: new Abstract: Unmixing reveals the spatial distribution and spectral details of different constituents, called endmembers, in a hyperspectral image. Because unmixing

researcharxiv-cs-cv
21 Apr 2026
Safety

IYKYK (But AI Doesn't): Automated Content Moderation Does Not Capture Communities' Heterogeneous Attitudes Towards Reclaimed Language

DGX agent

arXiv:2604.16654v1 Announce Type: new Abstract: Reclaimed slur usage is a common and meaningful practice online for many marginalized communities. It serves as a source of solidarity, identity, and sh

safetyarxiv-cs-cl
21 Apr 2026
Model Releases

Judge a Book by its Cover: Investigating Multi-Modal LLMs for Multi-Page Handwritten Document Transcription

DGX agent

arXiv:2502.20295v2 Announce Type: replace-cross Abstract: Handwriting text recognition (HTR) remains a challenging task. Existing approaches require fine-tuning on labeled data, which is impractical t

model-releasesarxiv-cs-cv
21 Apr 2026
Applications

Large Language Models Are Bad Dice Players: LLMs Struggle to Generate Random Numbers from Statistical Distributions

DGX agent

arXiv:2601.05414v2 Announce Type: replace Abstract: As large language models (LLMs) transition from chat interfaces to integral components of stochastic pipelines and systems approaching general intel

applicationsarxiv-cs-cl
21 Apr 2026
Local Ai

Learning to Seek Help: Dynamic Collaboration Between Small and Large Language Models

DGX agent

arXiv:2604.17827v1 Announce Type: new Abstract: Large language models (LLMs) offer strong capabilities but raise cost and privacy concerns, whereas small language models (SLMs) facilitate efficient an

local-aiarxiv-cs-cl
21 Apr 2026
Research

Linking Exteroception and Proprioception through Improved Contact Modeling for Soft Growing Robots

DGX agent

arXiv:2507.10694v2 Announce Type: replace Abstract: Passive deformation due to compliance is a commonly used benefit of soft robots, providing opportunities to achieve robust actuation with few active

researcharxiv-cs-ro
21 Apr 2026
Agents

Live LTL Progress Tracking: Towards Task-Based Exploration

DGX agent

arXiv:2604.17106v1 Announce Type: new Abstract: Motivated by the challenge presented by non-Markovian objectives in reinforcement learning (RL), we present a novel framework to track and represent the

agentsarxiv-cs-lg
21 Apr 2026
Research

LVLMs and Humans Ground Differently in Referential Communication

DGX agent

arXiv:2601.19792v3 Announce Type: replace Abstract: For generative AI agents to partner effectively with human users, the ability to accurately predict human intent is critical. But this ability to co

researcharxiv-cs-cl
21 Apr 2026
Agents

Matrix: Peer-to-Peer Multi-Agent Synthetic Data Generation Framework

DGX agent

arXiv:2511.21686v2 Announce Type: replace Abstract: Synthetic data has become increasingly important for training large language models, especially when real data is scarce, expensive, or privacy-sens

agentsarxiv-cs-cl
21 Apr 2026
Hardware

Muscle-inspired magnetic actuators that push, pull, crawl, and grasp

DGX agent

arXiv:2604.18090v1 Announce Type: new Abstract: Functional magnetic composites capable of large deformation, load bearing, and multifunctional motion are essential for next-generation adaptive soft ro

hardwarearxiv-cs-ro
21 Apr 2026
Safety

On the Convergence and Size Transferability of Continuous-depth Graph Neural Networks

DGX agent

arXiv:2510.03923v2 Announce Type: replace Abstract: Continuous-depth graph neural networks, also known as Graph Neural Differential Equations (GNDEs), combine the structural inductive bias of Graph Ne

safetyarxiv-cs-lg
21 Apr 2026
Model Releases

On the Predictive Power of Representation Dispersion in Language Models

DGX agent

arXiv:2506.24106v2 Announce Type: replace Abstract: We show that a language model's ability to predict text is tightly linked to the breadth of its embedding space: models that spread their contextual

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

PAC-Bayes Bounds for Gibbs Posteriors via Singular Learning Theory

DGX agent

arXiv:2604.17219v1 Announce Type: cross Abstract: We derive explicit non-asymptotic PAC-Bayes generalization bounds for Gibbs posteriors, that is, data-dependent distributions over model parameters ob

model-releasesarxiv-cs-lg
21 Apr 2026
Research

Pearmut: Human Evaluation of Translation Made Trivial

DGX agent

arXiv:2601.02933v3 Announce Type: replace Abstract: Human evaluation is the gold standard for multilingual NLP, but is often skipped in practice and substituted with automatic metrics because it is no

researcharxiv-cs-cl
21 Apr 2026
Model Releases

PersonalHomeBench: Evaluating Agents in Personalized Smart Homes

DGX agent

arXiv:2604.16813v1 Announce Type: cross Abstract: Agentic AI systems are rapidly advancing toward real-world applications, yet their readiness in complex and personalized environments remains insuffic

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

PFDelta: A Benchmark Dataset for Power Flow under Load, Generation, and Topology Variations

DGX agent

arXiv:2510.22048v3 Announce Type: replace Abstract: Power flow (PF) calculations are the backbone of real-time grid operations, across workflows such as contingency analysis (where repeated PF evaluat

model-releasesarxiv-cs-lg
21 Apr 2026
Applications

Physics-Informed Neural Networks for Biological 2D{+}t Reaction-Diffusion Systems

DGX agent

arXiv:2604.18548v1 Announce Type: new Abstract: Physics-informed neural networks (PINNs) provide a powerful framework for learning governing equations of dynamical systems from data. Biologically-info

applicationsarxiv-cs-lg
21 Apr 2026
Hardware

PiERN: Token-Level Routing for Integrating High-Precision Computation and Reasoning

DGX agent

arXiv:2509.18169v3 Announce Type: replace-cross Abstract: Tasks on complex systems require high-precision numerical computation to support decisions, but current large language models (LLMs) cannot in

hardwarearxiv-cs-cl
21 Apr 2026
← Previous
1…99100101102103…108
Next →