AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,428
  • Agents7,398
  • Applications5,301
  • Concepts5
  • Hardware1,785
  • Industry6,113
  • Local Ai4,833
  • Model Releases23,177
  • Research19,713
  • Safety13,092
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,428
  • Agents7,398
  • Applications5,301
  • Concepts5
  • Hardware1,785
  • Industry6,113
  • Local Ai4,833
  • Model Releases23,177
  • Research19,713
  • Safety13,092
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
86,428Total entries
1Added by human
86,427Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
50,764 results
Research

Physically Grounded Monocular Depth via Nanophotonic Wavefront Encoding

DGX agent

arXiv:2503.15770v3 Announce Type: replace-cross Abstract: Depth foundation models (DFMs) offer strong learned priors for 3D perception from single RGB images but lack physical depth cues, leading to a

researcharxiv-cs-cv
9 Jul 2026
Agents
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

RIMRULE: Improving Tool-Using Language Agents via MDL-Guided Rule Learning

DGX agent

arXiv:2601.00086v3 Announce Type: replace Abstract: Large language models (LLMs) often struggle to use tools reliably in domain-specific settings, where APIs may be idiosyncratic, under-documented, or

agentsarxiv-cs-cl
9 Jul 2026
Research

Seeing What Matters: Lesion-Aware High-Resolution Patch Discovery and Fusion for Chest X-ray Report Generation

DGX agent

arXiv:2607.06909v1 Announce Type: new Abstract: Despite rapid advances in chest X-ray (CXR) foundation models, most radiology report generation (RRG) systems still rely on heavily downsampled inputs (

researcharxiv-cs-cv
9 Jul 2026
Local Ai

Segmenting Low-Contrast XCTs of Concrete: An Unsupervised Approach

DGX agent

arXiv:2603.00127v2 Announce Type: replace Abstract: X-Ray Computed Tomography (XCT) is a compelling tool in experimental mechanics, capable of non-destructively extracting information pertaining to th

local-aiarxiv-cs-cv
9 Jul 2026
Research

SpiS-GAN: Spiral-Modulated Handwriting Synthesis with Star Operation

DGX agent

arXiv:2607.06949v1 Announce Type: new Abstract: Training robust handwriting recognition (HTR) systems requires massive amounts of annotated data, which is often difficult to acquire. While synthetic h

researcharxiv-cs-cv
9 Jul 2026
Safety

Trajectory Inference of Human Aging from Cross-Sectional DNA Methylation Data

DGX agent

arXiv:2607.06583v1 Announce Type: cross Abstract: DNA methylation (DNAm) serves as one of the most robust molecular biomarkers of biological aging. While conventional epigenetic clocks accurately pred

safetyarxiv-cs-lg
9 Jul 2026
Research

A Guiding Framework for K-12 Teachers in Creating AI-powered Learning Technologies through Vibe Coding

DGX agent

arXiv:2607.05406v1 Announce Type: cross Abstract: Large language models generate code from natural language prompts, enabling 'vibe coding,' which allows non-programmers to develop computational solut

researcharxiv-cs-ai
8 Jul 2026
Research

A Physics-Informed Neural Network Framework for Elastodynamic Wave Propagation in Bimaterial Systems

DGX agent

arXiv:2607.06479v1 Announce Type: new Abstract: Physics-informed neural networks (PINNs) provide a promising framework for solving partial differential equations while embedding the underlying physica

researcharxiv-cs-ai
8 Jul 2026
Research

AdaStop: Cost-Aware Early Stopping for DNN Test Selection

DGX agent

arXiv:2607.05461v1 Announce Type: cross Abstract: Existing methods for testing deep neural networks (DNNs) primarily prioritize test inputs likely to reveal model faults under a fixed labeling budget.

researcharxiv-cs-ai
8 Jul 2026
Research

Anti-Prompt: Image Protection against Text-Guided Image-to-Video Generation

DGX agent

arXiv:2607.01499v2 Announce Type: replace Abstract: Recent advances in Image-to-Video generation allow a single image to be animated into a convincing video under text guidance, raising serious copyri

researcharxiv-cs-cv
8 Jul 2026
Agents

CanvasAgent: Enabling Complex Image Creation and Editing via Visual Tool Orchestration

DGX agent

arXiv:2607.05465v1 Announce Type: cross Abstract: Complex image creation and editing often require more than a single generation or editing model. A user request may involve synthesizing images, local

agentsarxiv-cs-ai
8 Jul 2026
Agents

CurateEvo: Data-Curation Evolving for Agentic Post-Training

DGX agent

arXiv:2607.06140v1 Announce Type: new Abstract: Large language model (LLM) agents require post-training methods that can improve long-horizon decision making from environment feedback. However, existi

agentsarxiv-cs-cl
8 Jul 2026
Agents

Demonstrating TOFFEE: A Learned System for Synthesizing Data Agent Trajectories at Scale

DGX agent

arXiv:2607.06233v1 Announce Type: new Abstract: LLM-powered data agents are playing an increasingly important role in data-driven decision making. However, existing data agents struggle to generalize

agentsarxiv-cs-ai
8 Jul 2026
Research

Depression Symptoms and Relational Patterns in 187k ChatGPT Histories

DGX agent

arXiv:2607.05685v1 Announce Type: cross Abstract: Large language models are increasingly used as private, always-available conversational systems, but little is known about how people with depressive

researcharxiv-cs-ai
8 Jul 2026
Safety

FADRA: Frequency-Aware Diffusion with Residual Adaptation for Video Face Restoration

DGX agent

arXiv:2607.06389v1 Announce Type: new Abstract: Video face restoration (VFR) aims to recover high-quality and temporally consistent facial details from severely degraded video sequences; however, exis

safetyarxiv-cs-cv
8 Jul 2026
Research

FIELDS: Face reconstruction with accurate Inference of Expression using Learning with Direct Supervision

DGX agent

arXiv:2511.21245v3 Announce Type: replace Abstract: Monocular 3D face reconstruction estimates a 3D morphable model (3DMM) representation from a single image, providing geometry-aware expression codes

researcharxiv-cs-cv
8 Jul 2026
Research

FUSE: A Flow-based Mapping Between Shapes

DGX agent

arXiv:2511.13431v2 Announce Type: replace Abstract: We introduce a novel neural representation for maps between 3D shapes based on flow-matching models, which is computationally efficient and supports

researcharxiv-cs-cv
8 Jul 2026
Research

GraphBU: MILP Instance Generation with Graph-Native Block Units

DGX agent

arXiv:2607.06532v1 Announce Type: new Abstract: Mixed-integer linear programming (MILP) instances used for solver development are hard to obtain when models come from private or application-specific p

researcharxiv-cs-lg
8 Jul 2026
Applications

Image2Sim: Scaling Embodied Navigation via Generative Neural Simulator

DGX agent

arXiv:2607.05765v1 Announce Type: new Abstract: Embodied navigation aims to build agents that interpret multimodal goals, reason in 3D space, and reach target destinations reliably in the real world.

applicationsarxiv-cs-cv
8 Jul 2026
Safety

Information Gain-based Rollout Policy Optimization: An Adaptive Tree-Structured Rollout Approach for Multi-Turn LLM Agents

DGX agent

arXiv:2607.06223v1 Announce Type: new Abstract: Reinforcement learning has become a promising paradigm for improving large language model (LLM) agents on long-horizon search tasks, where the agent mus

safetyarxiv-cs-ai
8 Jul 2026
Safety

Is Domain Adaptation Always Helpful? A Frozen-Backbone Study of Cross-Domain Sentiment Transfer

DGX agent

arXiv:2607.05937v1 Announce Type: new Abstract: Sentiment analysis with frozen pre-trained language model (PLM) backbones has become a common paradigm, yet the practical benefit of explicit domain ada

safetyarxiv-cs-cl
8 Jul 2026
Research

LEGATO 2: Toward Multimodal Sheet Music Recognition and Understanding

DGX agent

arXiv:2607.05769v1 Announce Type: cross Abstract: We propose a novel pipeline, Legato 2, for extracting symbolic notation and semantic knowledge from images of sheet music. Legato 2 features the first

researcharxiv-cs-ai
8 Jul 2026
Hardware

Light-Omni: Reflex over Reasoning in Agentic Video Understanding with Long-Term Memory

DGX agent

arXiv:2607.05511v1 Announce Type: new Abstract: Agentic video understanding equips models with long-term memory to autonomously process and respond to continuous, long-horizon multimodal streams. Howe

hardwarearxiv-cs-cv
8 Jul 2026
Safety

Multimodal Analytics of Cybersecurity Crisis Preparation Exercises: What Predicts Success?

DGX agent

arXiv:2603.28553v2 Announce Type: replace-cross Abstract: Instructional alignment, the match between intended cognition and enacted activity, is central to effective instruction but hard to operationa

safetyarxiv-cs-lg
8 Jul 2026
Applications

Pitwall: Faithful Natural-Language Race-Strategy Briefings from a Calibrated Real-Time Monte Carlo Engine

DGX agent

arXiv:2607.06495v1 Announce Type: cross Abstract: Live sports commentary is grounded generation under a deadline: statements concern real, named athletes, the grounding state changes every few seconds

applicationsarxiv-cs-ai
8 Jul 2026
Agents

PolyJarvis: An LLM-Orchestrated Agent for Automated All-Atom Molecular Dynamics of Amorphous Homopolymers

DGX agent

arXiv:2604.02537v2 Announce Type: replace Abstract: All-atom molecular dynamics (MD) simulations can predict polymer properties from molecular structure, yet their execution requires specialized exper

agentsarxiv-cs-cl
8 Jul 2026
Agents

Prompt-to-Paper: Agentic AI System for Bioinformatics

DGX agent

arXiv:2607.05456v1 Announce Type: new Abstract: While recent advances in large language models have enabled end-to-end automated manuscript generation, existing systems suffer from three critical defi

agentsarxiv-cs-ai
8 Jul 2026
Local Ai

ProxyPose: 6-DoF Pose Tracking via Video-to-Video Translation

DGX agent

arXiv:2607.06555v1 Announce Type: new Abstract: Tracking the six-degree-of-freedom (6-DoF) pose of objects and surfaces from monocular video is a long-standing problem in computer vision. To tackle th

local-aiarxiv-cs-cv
8 Jul 2026
Research

Rethinking Visual Autoregressive Sampling with Information-Grounding Guidance

DGX agent

arXiv:2509.23876v3 Announce Type: replace-cross Abstract: Autoregressive (AR) models based on next-scale prediction have emerged as a powerful tool for image generation, but they face a critical weakn

researcharxiv-cs-ai
8 Jul 2026
Research

Segmentation before Answering: Pixel Grounding for MLLM Visual Reasoning

DGX agent

arXiv:2607.05798v1 Announce Type: cross Abstract: Recent advancements in Multimodal Large Language Models (MLLMs) have evolved from static perception to interleaved visual-language reasoning, often re

researcharxiv-cs-ai
8 Jul 2026
Research

SparseCtrl-HOI: Sparse Temporal Control for Human-Object Interaction Video Generation

DGX agent

arXiv:2607.05994v1 Announce Type: new Abstract: Human-Object Interaction (HOI) video generation aims to synthesize realistic videos of humans manipulating diverse objects, serving as a promising avenu

researcharxiv-cs-cv
8 Jul 2026
Agents

Strategic Bargaining in Multi-Buyer Markets: Reinforcement Learning from Verifiable Rewards for LLM Negotiations

DGX agent

arXiv:2607.05863v1 Announce Type: new Abstract: Negotiation is a fundamental strategic interaction in management science, characterized by agents attempting to reach agreements while protecting privat

agentsarxiv-cs-lg
8 Jul 2026
Safety

Superman: Unifying Skeleton and Vision for Human Motion Perception and Generation

DGX agent

arXiv:2602.02401v2 Announce Type: replace Abstract: Human motion analysis tasks, such as temporal 3D pose estimation, motion prediction, and motion in-betweening, play an essential role in computer vi

safetyarxiv-cs-cv
8 Jul 2026
Agents

TRIG: Trajectory-Rig Decoupled Metric Geometry Learning

DGX agent

arXiv:2607.05801v1 Announce Type: new Abstract: Vision-centric autonomous driving requires accurate metric geometry and ego-motion estimation from synchronized multi-camera observations. Recent visual

agentsarxiv-cs-cv
8 Jul 2026
Agents

TypeGo: An OS Runtime for Embodied Agents

DGX agent

arXiv:2607.05482v1 Announce Type: cross Abstract: Large language models (LLMs) can plan behavior for embodied agents from natural language, but treating the LLM as a request/response oracle on the cri

agentsarxiv-cs-ro
8 Jul 2026
Hardware

UBEP: Re-architecting Expert Parallelism Communication Library for Production Superpods

DGX agent

arXiv:2607.06202v1 Announce Type: cross Abstract: The deployment of Mixture-of-Experts (MoE) models on production high-bandwidth superpods, such as NVIDIA's NVL72/576 and Huawei's CloudMatrix384, intr

hardwarearxiv-cs-ai
8 Jul 2026
Safety

UniField: A Unified Field-Aware MRI Enhancement Framework

DGX agent

arXiv:2603.09223v2 Announce Type: replace Abstract: Magnetic Resonance Imaging (MRI) field-strength enhancement holds immense value for both clinical diagnostics and advanced research. However, existi

safetyarxiv-cs-cv
8 Jul 2026
Research

Universality of Benign Overfitting in Binary Linear Classification

DGX agent

arXiv:2501.10538v3 Announce Type: replace Abstract: The practical success of deep learning has led to the discovery of several surprising phenomena. One of these phenomena, that has spurred intense th

researcharxiv-cs-lg
8 Jul 2026
Agents

VaseMuseum: Digital Intelligent Museum for Ancient Greek Pottery

DGX agent

arXiv:2607.06374v1 Announce Type: new Abstract: Vision-language models (VLMs) have made interactive digital museums increasingly feasible by connecting 3D digitization with natural-language artifact e

agentsarxiv-cs-cv
8 Jul 2026
Agents

VASP Agent: An Agentic Framework for Autonomous First-principles Calculations

DGX agent

arXiv:2512.19458v2 Announce Type: replace Abstract: Large Language Models (LLMs) are increasingly embedded in agentic frameworks for scientific discovery. First-principles materials computation impose

agentsarxiv-cs-ai
8 Jul 2026
Safety

WordVoice: Explicit and Decoupled Multi-Dimensional Word-Level Control for LLM-Based TTS

DGX agent

arXiv:2607.06461v1 Announce Type: cross Abstract: While recent Large Language Model (LLM)-based Text-to-Speech (TTS) systems have achieved remarkable naturalness, they predominantly rely on implicit e

safetyarxiv-cs-cl
8 Jul 2026
Hardware

A Multi-Task Deep Learning Framework for Real-Time Intelligent Video Surveillance with Temporal Event Validation

DGX agent

arXiv:2607.03131v1 Announce Type: cross Abstract: Modern video surveillance systems generate far more video streams than human operators can effectively monitor, making automated analysis essential fo

hardwarearxiv-cs-ai
7 Jul 2026
Safety

ACPO: Adaptive Credit Policy Optimization via Fine-Grained Surrogate Entropy

DGX agent

arXiv:2607.03126v1 Announce Type: cross Abstract: Reinforcement Learning (RL) has substantially improved the reasoning ability of large language models (LLMs), but sparse outcome rewards still make to

safetyarxiv-cs-ai
7 Jul 2026
Applications

Active Learning on Adversarially Corrupted Graphs

DGX agent

arXiv:2607.04869v1 Announce Type: new Abstract: Motivated by real-world scenarios where malicious entities tamper with existing networks, we define a model where an adversary seeks to hide a set of co

applicationsarxiv-cs-lg
7 Jul 2026
Research

Adaptive Loss Balancing for Multi-Task Bioacoustic Classification of Bird Species and Call Types

DGX agent

arXiv:2607.03304v1 Announce Type: cross Abstract: Reliable analysis of bird vocalisations in passive acoustic monitoring requires models handling multiple, imbalanced annotation targets. We extend Bir

researcharxiv-cs-lg
7 Jul 2026
Local Ai

Agent Reinforcement Learning via Pivotal-Aware Self-Feedback Retry

DGX agent

arXiv:2607.03702v1 Announce Type: new Abstract: Large language model (LLM) agents have shown strong decision-making capabilities in long-horizon interactive tasks, yet they still struggle to effective

local-aiarxiv-cs-ai
7 Jul 2026
Safety

Agentic Artificial Intelligence for Multistage Physics Experiments at a Large-Scale User Facility Particle Accelerator

DGX agent

arXiv:2509.17255v2 Announce Type: replace-cross Abstract: We present the first language-model-driven agentic artificial intelligence (AI) system to autonomously execute multi-stage physics experiments

safetyarxiv-cs-ai
7 Jul 2026
Research

An Interpretable Deep Learning Framework for Discovery and Clinical Validation of Deep Radiomic Signatures in Tumor Classification

DGX agent

arXiv:2607.03593v1 Announce Type: cross Abstract: Imaging signatures are quantitative features extracted from medical images that provide clinically meaningful information for tumor diagnosis, charact

researcharxiv-cs-ai
7 Jul 2026
← Previous
1…759760761762763…1058
Next →