AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries90,259
  • Agents7,702
  • Applications5,506
  • Concepts5
  • Hardware1,896
  • Industry6,187
  • Local Ai5,048
  • Model Releases24,516
  • Research20,616
  • Safety13,635
  • Syntheses17
  • Tools1,677
  • Tutorials3,454

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries90,259
  • Agents7,702
  • Applications5,506
  • Concepts5
  • Hardware1,896
  • Industry6,187
  • Local Ai5,048
  • Model Releases24,516
  • Research20,616
  • Safety13,635
  • Syntheses17
  • Tools1,677
  • Tutorials3,454

Source
HumanDGX agent

Content type
90,259Total entries
1Added by human
90,258Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
64,475 results
Model Releases

Using Large Language Models to Support High Volume Application Review for an Undergraduate Research Program

DGX agent

arXiv:2606.05564v1 Announce Type: new Abstract: Undergraduate research programs such as the Summer Undergraduate Research Fellowship (SURF) at Purdue University receive thousands of applications every

model-releasesarxiv-cs-cl
5 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Using street view images and visual LLMs to predict heritage values for governance support: Risks, ethics, and policy implications

DGX agent

arXiv:2601.06056v2 Announce Type: replace-cross Abstract: During 2025 and 2026, the Energy Performance of Buildings Directive is being implemented in the European Union member states, requiring all me

safetyarxiv-cs-cv
5 Jun 2026
Model Releases

V2V-Bench: A Comprehensive Benchmark for Video-to-Video Generation Evaluation

DGX agent

arXiv:2606.05665v1 Announce Type: new Abstract: Video-to-video (V2V) generation is difficult to evaluate because outputs must both follow editing instructions and preserve frame-level correspondence w

model-releasesarxiv-cs-cv
5 Jun 2026
Safety

Value-and-Structure Alignment for Routing-Consistent Quantization of Mixture-of-Experts Models

DGX agent

arXiv:2606.05688v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) models scale foundation models efficiently by activating only a subset of experts for each token, but their large number of exp

safetyarxiv-cs-cl
5 Jun 2026
Local Ai

VASO: Formally Verifiable Self-Evolving Skills for Physical AI Agents

DGX agent

arXiv:2606.05395v1 Announce Type: new Abstract: Reusable robot skills are becoming the basic units through which embodied agents turn open-ended instructions into long-horizon physical behavior. We ar

local-aiarxiv-cs-ro
5 Jun 2026
Research

Vavanagi: a Community-run Platform for Documentation of the Hula Language in Papua New Guinea

DGX agent

arXiv:2603.14210v2 Announce Type: replace Abstract: We present Vavanagi, a community-run platform for Hula (Vula'a), an Austronesian language of Papua New Guinea with approximately 10,000 speakers. Va

researcharxiv-cs-cl
5 Jun 2026
Safety

ViCuR: Visual Cues as Recoverable Privilege for Multimodal On-Policy Distillation

DGX agent

arXiv:2606.05718v1 Announce Type: new Abstract: On-policy distillation (OPD) improves reasoning by training a student on trajectories sampled from its own policy under supervision from a teacher. In m

safetyarxiv-cs-cv
5 Jun 2026
Model Releases

Video-Rate Streaming Stylization on a Vision-Aware MLLM-Conditioned Edit Diffusion: Asymmetric Batched Inference on a Distilled UNet + MLLM Text Encoder

DGX agent

arXiv:2606.05981v1 Announce Type: new Abstract: Aggressive distillation of the diffusion U-Net inverts the per-frame bottleneck of real-time text-to-image pipelines: once the denoiser is a 4-step or 1

model-releasesarxiv-cs-cv
5 Jun 2026
Model Releases

VideoKR: Towards Knowledge- and Reasoning-Intensive Video Understanding

DGX agent

arXiv:2606.05259v1 Announce Type: new Abstract: We introduce VideoKR, the first large-scale training corpus specifically designed to strengthen knowledge- and reasoning-intensive video understanding.

model-releasesarxiv-cs-cv
5 Jun 2026
Local Ai

Vision Hopfield Memory Networks

DGX agent

arXiv:2603.25157v2 Announce Type: replace-cross Abstract: Recent vision and multimodal foundation backbones, such as Transformer families and state-space models like Mamba, have achieved remarkable pr

local-aiarxiv-cs-cv
5 Jun 2026
Research

Visual Commonsense Driven Knowledge Refinements for Scene Graph Generation

DGX agent

arXiv:2606.06369v1 Announce Type: new Abstract: Learning-driven Scene Graph Generation (SGG) models excel on frequent relation types but degrade sharply under annotation sparsity, failing to capture r

researcharxiv-cs-cv
5 Jun 2026
Agents

Visuotactile and Explicitly Force-Controlled Robotic Ultrasound for Abdominal Volumetric Reconstruction

DGX agent

arXiv:2606.05848v1 Announce Type: new Abstract: In this paper, we present a robotic ultrasound acquisition system that integrates stereo vision, touch-based feedback, and expert-informed strategies to

agentsarxiv-cs-ro
5 Jun 2026
Safety

VOLD: Reasoning Transfer from LLMs to Vision-Language Models via On-Policy Distillation

DGX agent

arXiv:2510.23497v3 Announce Type: replace Abstract: Training vision-language models (VLMs) for complex reasoning remains a challenging task, i.a. due to the scarcity of high-quality image-text reasoni

safetyarxiv-cs-cv
5 Jun 2026
Safety

VOLT: Vision and Language Trajectory Segmentation for Faster-than-Demonstration Policies

DGX agent

arXiv:2606.06323v1 Announce Type: new Abstract: Humans often take longer to demonstrate a task than a robot would need to execute it. Rather than learning to replicate the demonstration at the same pa

safetyarxiv-cs-ro
5 Jun 2026
Model Releases

VTI-CoT: Visual-Textual Interleaved Chain of Thought for Video Reasoning

DGX agent

arXiv:2606.05736v1 Announce Type: new Abstract: Video reasoning aims to understand complex temporal events and causal relationships within videos. Recently, Chain-of-Thought (CoT) has been introduced

model-releasesarxiv-cs-cv
5 Jun 2026
Model Releases

VZCrash: A Large-Scale IMU Dataset of Ego-Vehicle Crashes

DGX agent

arXiv:2606.06074v1 Announce Type: new Abstract: We introduce VZCrash, the largest publicly available dataset of real-world vehicle collision data featuring Inertial Measurement Unit (IMU) telemetry. T

model-releasesarxiv-cs-cv
5 Jun 2026
Research

Wave Focusing in Metamaterials: Tactile Displays Beyond the Diffraction Limit

DGX agent

arXiv:2606.05572v1 Announce Type: cross Abstract: We address the challenge of engineering distributed haptic displays capable of reproducing multiple localized, independently addressable vibrations --

researcharxiv-cs-ro
5 Jun 2026
Model Releases

Waypoints Matter: A Systematic Study for Sampling-Based Trajectory Planning

DGX agent

arXiv:2606.06366v1 Announce Type: new Abstract: Real-time autonomous driving commonly relies on sampling-based trajectory planners that link candidate trajectories to target waypoints along the road c

model-releasesarxiv-cs-ro
5 Jun 2026
Research

What Makes Two Language Models Think Alike?

DGX agent

arXiv:2406.12620v3 Announce Type: replace Abstract: Do architectural and training differences influence the way models represent and process language? Traditional similarity metrics tell us whether tw

researcharxiv-cs-cl
5 Jun 2026
Research

What Objects Enable, Not What They Are: Functional Latent Spaces for Affordance Reasoning

DGX agent

arXiv:2606.05533v1 Announce Type: cross Abstract: Existing robot planning systems rely on appearance-based reasoning, where visual observations are encoded into latent spaces organized around object a

researcharxiv-cs-cv
5 Jun 2026
Safety

What's in a Name? Morphological Shortcuts by LLMs in Pharmacology

DGX agent

arXiv:2606.05616v1 Announce Type: new Abstract: The morphological form of a word can often give cues to its meaning, but purely relying on these mappings can lead to overgeneralization in high-stakes

safetyarxiv-cs-cl
5 Jun 2026
Applications

What's Under the Skin? Estimating Swine Body Condition

DGX agent

arXiv:2606.05611v1 Announce Type: new Abstract: Sow body condition is an important indicator for growers as it has a large impact on lactation performance and piglet survival. However, body condition

applicationsarxiv-cs-cv
5 Jun 2026
Safety

When AI Says It Feels

DGX agent

arXiv:2606.05734v1 Announce Type: cross Abstract: Large language models (LLMs) are generally constrained from expressing feelings through human-preference alignment in post-training processes. This po

safetyarxiv-cs-cl
5 Jun 2026
Safety

When Evidence is Sparse: Weakly Supervised Early Failure Alerting in Dialogs and LLM-Agent Trajectories

DGX agent

arXiv:2606.05414v1 Announce Type: new Abstract: Early failure alerting requires deciding, while a dialog or agent trajectory is still unfolding, whether to flag it as likely to fail. This is challengi

safetyarxiv-cs-cl
5 Jun 2026
Research

When New Generators Arrive: Lifelong Machine-Generated Text Attribution via Ridge Feature Transfer

DGX agent

arXiv:2606.05626v1 Announce Type: new Abstract: Machine-generated text (MGT) attribution aims to identify the specific generator responsible for a given text, thereby providing fine-grained evidence f

researcharxiv-cs-cl
5 Jun 2026
Research

Where does Absolute Position come from in decoder-only Transformers?

DGX agent

arXiv:2606.06160v1 Announce Type: cross Abstract: RoPE-trained transformers distinguish absolute position in their attention patterns, even though RoPE encodes only relative offsets in the inner produ

researcharxiv-cs-cl
5 Jun 2026
Local Ai

Where, What, Why, and Importance: Structured Defect Grounding for Text-to-Image Feedback

DGX agent

arXiv:2606.06113v1 Announce Type: new Abstract: Despite generating increasingly photorealistic images, text-to-image (T2I) models still exhibit localized, subtle, and structurally complex failures. Di

local-aiarxiv-cs-cv
5 Jun 2026
Hardware

World-Language-Action Model for Unified World Modeling, Language Reasoning, and Action Synthesis

DGX agent

arXiv:2606.05979v1 Announce Type: new Abstract: We propose world-language-action (WLA) models as a new class of embodied foundation models. WLA takes textual instructions, images, and robot states as

hardwarearxiv-cs-ro
5 Jun 2026
Model Releases

Would you still call this Dax? Novel Visual References in VLMs and Humans

DGX agent

arXiv:2606.05409v1 Announce Type: cross Abstract: Vision-language models (VLMs), like human learners, are frequently exposed to new visual concepts, but how they map novel visual references to languag

model-releasesarxiv-cs-cl
5 Jun 2026
Research

You Only Index Once: Cross-Layer Sparse Attention with Shared Routing

DGX agent

arXiv:2606.06467v1 Announce Type: new Abstract: Long-context inference in modern LLMs is increasingly constrained by decoding efficiency, especially in reasoning-heavy settings where models generate l

researcharxiv-cs-cl
5 Jun 2026
Model Releases

YouZhi: Towards High-Concurrency Financial LLMs via Adaptive GQA-to-MLA Transition

DGX agent

arXiv:2606.05868v1 Announce Type: new Abstract: Large language models (LLMs) drive significant financial innovations, yet their high-concurrency deployment is severely bottlenecked by KV cache memory

model-releasesarxiv-cs-cl
5 Jun 2026
Model Releases

100-LongBench: Are de facto Long-Context Benchmarks Literally Evaluating Long-Context Ability?

DGX agent

arXiv:2505.19293v2 Announce Type: replace-cross Abstract: Long-context capability is considered one of the most important abilities of LLMs, as a truly long context-capable LLM enables users to effort

model-releasesarxiv-cs-ai
4 Jun 2026
Research

3D Temporal Analysis for Autism Spectrum Disorder Screening During Attention Tasks

DGX agent

arXiv:2606.04836v1 Announce Type: new Abstract: Accurate Autism Spectrum Disorder (ASD) screening for school-age children is crucial to identify cases that may have been missed earlier and to enable t

researcharxiv-cs-cv
4 Jun 2026
Applications

3DThinkVLA: Endowing Vision-Language-Action Models with Latent 3D Priors via 3D-Thinking-Guided Co-training

DGX agent

arXiv:2606.04436v1 Announce Type: new Abstract: We propose a 3D-thinking-guided co-training framework that enables vision-language-action (VLA) models to perform 3D spatial reasoning implicitly during

applicationsarxiv-cs-cv
4 Jun 2026
Safety

3PoinTr: 3D Point Tracks for Learning Manipulation from Unconstrained Human Videos

DGX agent

arXiv:2603.08485v2 Announce Type: replace Abstract: Learning manipulation policies from human videos could greatly reduce the need for expensive robot demonstrations, but existing approaches typically

safetyarxiv-cs-ro
4 Jun 2026
Applications

4D Reconstruction from Sparse Dynamic Cameras

DGX agent

arXiv:2606.04593v1 Announce Type: new Abstract: Although dynamic 3D (i.e., 4D) reconstruction from a monocular dynamic camera has recently advanced, it remains fundamentally limited by depth ambiguity

applicationsarxiv-cs-cv
4 Jun 2026
Model Releases

A Cookbook of 3D Vision: Data, Learning Paradigms, and Application

DGX agent

arXiv:2606.04291v1 Announce Type: new Abstract: 3D vision has rapidly evolved, driven by increasingly diverse data representations, learning paradigms, and modeling strategies. Yet the field remains f

model-releasesarxiv-cs-cv
4 Jun 2026
Research

A French Corpus Annotated for Multiword Expressions with Adverbial Function

DGX agent

arXiv:2606.04828v1 Announce Type: new Abstract: This paper presents a French corpus annotated for multiword expressions (MWEs) with adverbial function. This corpus is designed for investigation on inf

researcharxiv-cs-cl
4 Jun 2026
Research

A General Framework for Dynamic Consistent Submodular Maximization

DGX agent

arXiv:2606.04946v1 Announce Type: cross Abstract: Consistency is an important property in dynamic submodular maximization and entails maintaining a near-optimal solution at all times, making only a sm

researcharxiv-cs-lg
4 Jun 2026
Local Ai

A Geometric Characterization of the Stationary Plateau for Two-Layer Neural Networks

DGX agent

arXiv:2606.04327v1 Announce Type: cross Abstract: We investigate the geometric structure of stationary plateaus that arise in the loss landscape of two-layer neural networks with smooth activation fun

local-aiarxiv-cs-ai
4 Jun 2026
Local Ai

A Geometric View of Counterfactual Behavior: Interaction of Boundary Proximity and Local Support

DGX agent

arXiv:2606.04209v1 Announce Type: new Abstract: Counterfactual explanations seek small, semantically meaningful changes to an input that alter a model's prediction, and are widely used to interpret an

local-aiarxiv-cs-lg
4 Jun 2026
Safety

A Goal-Set Characterization of Task Composition in the Boolean Task Algebra

DGX agent

arXiv:2606.04053v1 Announce Type: cross Abstract: The Boolean Task Algebra (BTA) provides a principled framework for zero-shot task composition in reinforcement learning by equipping goal-reaching tas

safetyarxiv-cs-ai
4 Jun 2026
Research

A Latent Variable Framework for Scaling Laws in Large Language Models

DGX agent

arXiv:2512.06553v2 Announce Type: replace-cross Abstract: We propose a statistical framework built on latent variable modeling for scaling laws of large language models (LLMs). Our work is motivated b

researcharxiv-cs-lg
4 Jun 2026
Model Releases

A New Angle on Bones: Robust Pose Estimation in X-Ray and Ultrasound

DGX agent

arXiv:2606.04700v1 Announce Type: new Abstract: Measuring the angle between bone structures is a routine task in medical image analysis and provides a key quantitative parameter for diagnosis and trea

model-releasesarxiv-cs-cv
4 Jun 2026
Research

A Normative Intermediate Representation for ASP-Based Compliance Reasoning

DGX agent

arXiv:2606.04619v1 Announce Type: new Abstract: We propose MONIR, a Modalized-Output Normative Intermediate Representation for ASP-based compliance reasoning. Its core fragment has a staged operationa

researcharxiv-cs-ai
4 Jun 2026
Safety

A Pathology Foundation Model for Gastric Cancer with Real-World Validation

DGX agent

arXiv:2606.04792v1 Announce Type: new Abstract: Gastric cancer remains a major cause of cancer mortality, yet its histological and molecular heterogeneity complicates diagnosis and risk stratification

safetyarxiv-cs-cv
4 Jun 2026
Model Releases

A Study of the Scale Invariant Signal to Distortion Ratio in Speech Separation with Noisy References

DGX agent

arXiv:2508.14623v2 Announce Type: replace-cross Abstract: This paper examines the implications of using the Scale-Invariant Signal-to-Distortion Ratio (SI-SDR) as both evaluation and training objectiv

model-releasesarxiv-cs-ai
4 Jun 2026
Research

A Systematic Analysis of Linguistic Features in AI-Generated Text Detection Across Domains and Models

DGX agent

arXiv:2606.04177v1 Announce Type: cross Abstract: Interpretable linguistic features offer a promising approach for explaining why a given text appears machine-generated, particularly for non-expert us

researcharxiv-cs-ai
4 Jun 2026
← Previous
1…630631632633634…1344
Next →