AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,356
  • Agents7,554
  • Applications5,409
  • Concepts5
  • Hardware1,835
  • Industry6,165
  • Local Ai4,929
  • Model Releases23,869
  • Research20,124
  • Safety13,369
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,356
  • Agents7,554
  • Applications5,409
  • Concepts5
  • Hardware1,835
  • Industry6,165
  • Local Ai4,929
  • Model Releases23,869
  • Research20,124
  • Safety13,369
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
Human
88,356Total entries
1Added by human
88,355Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
62,897 results
Model Releases

Towards Fast and Effective Long Video Understanding of Multimodal Large Language Models via Adaptive Quasi-Gaussian Sampling

DGX agent

arXiv:2606.24187v1 Announce Type: new Abstract: Long video understanding remains a daunting challenge for Multimodal Large Language Models (MLLMs) due to the excessive computation and memory footprint

model-releasesarxiv-cs-cv
24 Jun 2026
DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Applications

Towards Federated Long-Tailed Graph Learning: An Energy-Guided Dual Decoupling Approach

DGX agent

arXiv:2606.24237v1 Announce Type: new Abstract: Federated Graph Learning facilitates collaborative graph modeling across distributed clients while preserving data privacy. However, real-world data cat

applicationsarxiv-cs-ai
24 Jun 2026
Model Releases

Towards Spec Learning: Inference-Time Alignment from Preference Pairs

DGX agent

arXiv:2606.24004v1 Announce Type: cross Abstract: Steering a large language model (LLM) toward a desired behavior typically relies on an iterative process of hand-crafting a prompt based on a careful

model-releasesarxiv-cs-ai
24 Jun 2026
Research

Towards Version-aware Operations and Transaction Memories for Multi-layer MeMo

DGX agent

arXiv:2606.24040v1 Announce Type: cross Abstract: MeMo proposes language models with explicit multi-layer correlation matrix memories (CMMs), where memorization, retrieval, and forgetting are architec

researcharxiv-cs-ai
24 Jun 2026
Research

Tractable Reasoning and Conjunctive Query Answering for Defeasible DL-Lite under Rational Closure

DGX agent

arXiv:2606.24279v1 Announce Type: new Abstract: In Description Logics (DLs), reasoning under Rational Closure (RC) is a well-known and widely accepted non-monotonic formalism to handle defeasible know

researcharxiv-cs-ai
24 Jun 2026
Research

Training-free Cross-domain Few-shot Segmentation via Robust Semantic Representation and Matching

DGX agent

arXiv:2606.24297v1 Announce Type: new Abstract: Cross-domain Few-shot Segmentation (CD-FSS) aims to transfer knowledge learned from source domain to distinct target domains, segmenting unseen target c

researcharxiv-cs-cv
24 Jun 2026
Research

Transformation Behavior of Images in Latent Space

DGX agent

arXiv:2606.24430v1 Announce Type: cross Abstract: Training of neural networks for histopathology classification tasks typically relies on data encoding into latent space, which reduces complexity and

researcharxiv-cs-ai
24 Jun 2026
Model Releases

Transformer-Based Language Models Across Domain Verticals: Architectures, Applications and Critical Assessment

DGX agent

arXiv:2606.24331v1 Announce Type: new Abstract: Transformer-based language models have become the default substrate for natural language processing and the pace of new releases has made it hard for pr

model-releasesarxiv-cs-cl
24 Jun 2026
Model Releases

Tri-Efficient Transfer Learning for Point Cloud Videos

DGX agent

arXiv:2606.24175v1 Announce Type: new Abstract: While point cloud foundation models have significantly advanced point cloud video understanding, existing parameter-efficient fine-tuning (PEFT) methods

model-releasesarxiv-cs-cv
24 Jun 2026
Model Releases

Trimming the Long-Tail of Visual World Modeling Evaluation

DGX agent

arXiv:2606.24256v1 Announce Type: new Abstract: Physical interactions follow a long-tailed distribution: a set of common and regular interactions dominates human experience and visual data, while a br

model-releasesarxiv-cs-cv
24 Jun 2026
Model Releases

TrOCR for Medieval HTR: A Systematic Ablation Study with Cross-Dataset Validation

DGX agent

arXiv:2606.24302v1 Announce Type: new Abstract: Fine-tuning transformer-based handwritten text recognition (HTR) models on medieval manuscripts is challenging because these models are pre-trained on m

model-releasesarxiv-cs-cv
24 Jun 2026
Safety

TruncProof: A Guardrail for LLM-based JSON Generation under Token-Length Constraints

DGX agent

arXiv:2605.13076v2 Announce Type: replace Abstract: The LLM-based generation of machine-readable outputs such as JSON has attracted significant attention for integration with external systems. However

safetyarxiv-cs-cl
24 Jun 2026
Research

Trustworthy Image Authentication using Forensic Knowledge Graphs

DGX agent

arXiv:2606.23917v1 Announce Type: new Abstract: Advances in generative AI have made image falsification highly realistic, demanding trustworthy authentication systems. Existing forensic detectors can

researcharxiv-cs-cv
24 Jun 2026
Model Releases

Tuning without Peeking: Provable Generalization Bounds and Robust LLM Post-Training

DGX agent

arXiv:2507.01752v4 Announce Type: replace-cross Abstract: Gradient-based optimization is the workhorse of deep learning, offering efficient and scalable training via backpropagation. However, exposing

model-releasesarxiv-cs-ai
24 Jun 2026
Hardware

TurboMPC: Fast, Scalable, and Differentiable Model Predictive Control on the GPU

DGX agent

arXiv:2606.24039v1 Announce Type: new Abstract: Robotics increasingly relies on GPUs for parallel simulation, large-scale learning, and neural-network inference. For model predictive control (MPC) to

hardwarearxiv-cs-ro
24 Jun 2026
Applications

TuringViT: Making SOTA Vision Transformers Accessible to All

DGX agent

arXiv:2606.24253v1 Announce Type: new Abstract: Modern VLMs and VLA systems commonly adopt off-the-shelf ViTs such as SigLIP2 as visual encoders, but diverse downstream requirements in latency, tempor

applicationsarxiv-cs-cv
24 Jun 2026
Research

Uncertainty-Aware Longitudinal Forecasting of Alzheimer's Disease Progression Using Deep Learning

DGX agent

arXiv:2606.24604v1 Announce Type: new Abstract: Longitudinal modelling of Alzheimer's disease progression is clinically useful only if it can describe not just the most likely next diagnosis, but how

researcharxiv-cs-ai
24 Jun 2026
Research

Understanding Deep Representation Learning via Layerwise Feature Compression and Discrimination

DGX agent

arXiv:2311.02960v5 Announce Type: replace-cross Abstract: Over the past decade, deep learning has proven to be a highly effective tool for learning meaningful features from raw data. However, it remai

researcharxiv-cs-cv
24 Jun 2026
Model Releases

UniDrive: A Unified Vision-Language and Grounding Framework for Interpretable Risk Understanding in Autonomous Driving

DGX agent

arXiv:2606.24759v1 Announce Type: cross Abstract: Recent multimodal large language models (MLLMs) have shown strong potential for autonomous driving scene understanding, yet existing methods still fac

model-releasesarxiv-cs-ai
24 Jun 2026
Research

Uniform Sampling from High-dimensional Spectral Norm Balls

DGX agent

arXiv:2606.24134v1 Announce Type: cross Abstract: Motivated by an application in machine learning optimization, this paper focuses on the challenges of sampling a matrix uniformly from the unit spectr

researcharxiv-cs-lg
24 Jun 2026
Model Releases

UniRED: Unified RGB-D Video Frame Interpolation with Event Guidance

DGX agent

arXiv:2606.24282v1 Announce Type: new Abstract: High frame-rate RGB-D videos are crucial for a variety of downstream tasks, including motion analysis, dynamic scene understanding, and 3D reconstructio

model-releasesarxiv-cs-cv
24 Jun 2026
Safety

UniTranslator: A Unified Multi-modal Framework for End-to-end In-Image Machine Translation

DGX agent

arXiv:2606.24333v1 Announce Type: new Abstract: In-Image Machine Translation (IIMT) aims to translate scene text in an image and render the translated text back into the original regions while preserv

safetyarxiv-cs-cv
24 Jun 2026
Agents

Universal Guideline-Driven Image Clustering via a Hybrid LLM Agent

DGX agent

arXiv:2606.24094v1 Announce Type: new Abstract: Unifying image clustering across different clustering scenarios remains challenging due to fundamental gaps among tasks. We introduce a Guideline-Driven

agentsarxiv-cs-cv
24 Jun 2026
Safety

UOL@IDEM at BEA 2026 Shared Task 1: Neural Fusion and Feature-Rich Modeling for L1-Aware Vocabulary Difficulty Prediction

DGX agent

arXiv:2606.24501v1 Announce Type: new Abstract: This paper describes UOL@IDEM's closed-track submission to the BEA 2026 shared task on L1-aware vocabulary difficulty prediction. We model the task as r

safetyarxiv-cs-cl
24 Jun 2026
Model Releases

Variational Model Merging for Pareto Front Estimation in Multitask Finetuning

DGX agent

arXiv:2412.08147v2 Announce Type: replace-cross Abstract: Pareto fronts are useful to find good task-mixing strategies for multitask finetuning, but they are also costly to compute. To reduce costs, r

model-releasesarxiv-cs-ai
24 Jun 2026
Agents

Varying Bundle Size Reactive Multi-Task Assignment using Selective Cost Estimation for Multi-Agent Systems

DGX agent

arXiv:2606.24462v1 Announce Type: new Abstract: This paper presents a scalable framework for multi-robot task allocation in complex environments where estimating task execution costs is computationall

agentsarxiv-cs-ro
24 Jun 2026
Safety

Verifiable Foundation Models for Robot Safety

DGX agent

arXiv:2606.23754v1 Announce Type: cross Abstract: Deploying foundation models for robot control raises a central challenge: the expressive power that enables rich, multimodal perception also makes the

safetyarxiv-cs-lg
24 Jun 2026
Model Releases

VeriPilot: An LLM-Powered Verilog Debugging Framework

DGX agent

arXiv:2606.23759v1 Announce Type: cross Abstract: Verilog debugging remains one of the most time-consuming stages in digital circuit design. Recent advances in Large Language Models (LLMs) have enable

model-releasesarxiv-cs-ai
24 Jun 2026
Research

VeryTrace: Verifying Reasoning Traces through Compilable Formalism and Structured Verification

DGX agent

arXiv:2606.24124v1 Announce Type: new Abstract: Multi-step reasoning with Chain-of-Thought (CoT) prompting remains fragile: logical errors or hallucinations in early steps silently propagate, producin

researcharxiv-cs-ai
24 Jun 2026
Model Releases

video-SALMONN-R^3: Learning to ReWatch, ReAsk, and ReAnswer for Efficient Video Understanding

DGX agent

arXiv:2606.24477v1 Announce Type: cross Abstract: Video large language models (LLMs) are often constrained by computation and memory budgets, leading them to use reduced frame rates and spatial resolu

model-releasesarxiv-cs-ai
24 Jun 2026
Research

VieSpeaker: A Large-Scale Vietnamese Speaker Recognition Dataset Beyond Visual Dependency

DGX agent

arXiv:2606.24066v1 Announce Type: cross Abstract: Speaker recognition has advanced rapidly with large-scale training datasets, yet Vietnamese remains under-resourced, with existing corpora limited in

researcharxiv-cs-cl
24 Jun 2026
Applications

VisChronos: Revolutionizing Image Captioning Through Real-Life Events

DGX agent

arXiv:2606.24058v1 Announce Type: new Abstract: This paper aims to bridge the semantic gap between visual content and natural language understanding by leveraging historical events in the real world a

applicationsarxiv-cs-cv
24 Jun 2026
Model Releases

VisCritic: Visual State Comparison as Process Reward for GUI Agents

DGX agent

arXiv:2606.24525v1 Announce Type: new Abstract: GUI agents powered by vision-language models show strong potential for automating digital tasks, yet frequently fail in long-horizon scenarios due to th

model-releasesarxiv-cs-cv
24 Jun 2026
Agents

Vision-Language Model Reasoning for Contextual Semantic Mapping in Intralogistics

DGX agent

arXiv:2606.24814v1 Announce Type: new Abstract: Autonomous mobile robots operating in intralogistics environments rely on geometric maps for localization and navigation, but lack semantic understandin

agentsarxiv-cs-ro
24 Jun 2026
Local Ai

VistaRef: Boosting Visual Spatial Orientation Awareness for Pointing-to-Object Detection

DGX agent

arXiv:2606.24498v1 Announce Type: new Abstract: Grounding deictic gestures in natural images is fundamental to AR and human-robot collaboration, providing a basis for seamless spatial interaction. Whi

local-aiarxiv-cs-cv
24 Jun 2026
Research

Visualizing 'We the People': Bridging the Perception Gap through Pluralistic Data Storytelling

DGX agent

arXiv:2606.24635v1 Announce Type: cross Abstract: Traditional visual data storytelling relies on binary graphics that depict two simplified groups in conflict. This can increase political polarization

researcharxiv-cs-ai
24 Jun 2026
Applications

ViTexQA: A Multi-Frame Temporal Perception Dataset for Video Text Question Answering

DGX agent

arXiv:2606.24602v1 Announce Type: new Abstract: Despite remarkable progress in multimodal understanding, current MLLMs still exhibit limitations in video text understanding, particularly when semantic

applicationsarxiv-cs-cv
24 Jun 2026
Hardware

VoltanaLLM: Energy-Efficient and SLO-Aware Disaggregated LLM Serving via Adaptive Frequency Control and State-Space Routing

DGX agent

arXiv:2509.04827v3 Announce Type: replace-cross Abstract: The energy cost of Large Language Model (LLM) inference is rapidly becoming a barrier to sustainable and scalable deployment. Although modern

hardwarearxiv-cs-ai
24 Jun 2026
Research

VSANet: View-aware Sparse Attention Network for Light Field Image Denoising

DGX agent

arXiv:2606.24737v1 Announce Type: new Abstract: Light field (LF) image denoising is challenging due to the high-dimensional structure of LF data. While noise is independent across sub-aperture images,

researcharxiv-cs-cv
24 Jun 2026
Research

Weight-Space Geometry of Offline Reasoning Training

DGX agent

arXiv:2606.23740v1 Announce Type: cross Abstract: Offline reinforcement-learning losses (RFT, RIFT, DFT, Offline GRPO, DPO) are widely used to distill reasoning from large teachers into smaller studen

researcharxiv-cs-ai
24 Jun 2026
Safety

What Do Flow-Based Inverse Solvers Approximate? A Posterior-Transport View

DGX agent

arXiv:2606.24516v1 Announce Type: new Abstract: A growing family of training-free solvers -- FlowDPS, FLOWER, PnP-Flow and their diffusion ancestors (DPS, DAPS) -- repurpose a pretrained flow-matching

safetyarxiv-cs-cv
24 Jun 2026
Model Releases

What Does ODRL Mean? A Cross-Level Ontological Grounding of Permissions, Prohibitions, and Duties in UFO-L

DGX agent

arXiv:2606.24344v1 Announce Type: cross Abstract: ODRL policy evaluators produce verdicts, but say nothing about the normative positions a policy brings into existence, the authority structures those

model-releasesarxiv-cs-ai
24 Jun 2026
Research

What's Missing in Vision-Language Models? Probing Their Struggles with Causal Order Reasoning

DGX agent

arXiv:2506.00869v3 Announce Type: replace Abstract: Despite the impressive performance of vision-language models (VLMs) on downstream tasks, their ability to understand and reason about causal relatio

researcharxiv-cs-cl
24 Jun 2026
Safety

When AI Meets Finance (StockAgent): Large Language Model-based Stock Trading in Simulated Real-world Environments

DGX agent

arXiv:2407.18957v5 Announce Type: replace-cross Abstract: Can AI Agents simulate real-world trading environments to investigate the impact of external factors on stock trading activities (e.g., macroe

safetyarxiv-cs-ai
24 Jun 2026
Safety

When CQs Go Wrong: Challenges in CQ Verification with OE-Assist

DGX agent

arXiv:2606.24619v1 Announce Type: new Abstract: Competency Questions (CQs) are the central component of CQ-verification, an established process in which an ontology is evaluated against a set of natur

safetyarxiv-cs-ai
24 Jun 2026
Model Releases

When Helpfulness Overrides Causal Caution: Context-Dependent Suppression and Recovery in LLMs

DGX agent

arXiv:2606.24370v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly integrated into decision-support roles in business and policy contexts. While prior benchmark studies have

model-releasesarxiv-cs-ai
24 Jun 2026
Safety

When Preferences Fail to Become Incentives: A Utility-Behavior Gap in Large Language Models

DGX agent

arXiv:2606.22974v2 Announce Type: replace Abstract: Recent work on preference elicitation in large language models (LLMs) has demonstrated that, when given a series of choices between two outcomes, LL

safetyarxiv-cs-ai
24 Jun 2026
Model Releases

When Retrieval Metrics Mislead: Measuring Policy Signal in Long-Horizon Tool-Use Agents

DGX agent

arXiv:2606.23937v1 Announce Type: cross Abstract: Exact-match retrieval recall is often used as a proxy for whether a retriever supplies useful policy context to a downstream decision model. We test t

model-releasesarxiv-cs-ai
24 Jun 2026
← Previous
1…488489490491492…1311
Next →