AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,042
  • Agents7,457
  • Applications5,325
  • Concepts5
  • Hardware1,802
  • Industry6,143
  • Local Ai4,863
  • Model Releases23,393
  • Research19,837
  • Safety13,180
  • Syntheses17
  • Tools1,671
  • Tutorials3,349

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,042
  • Agents7,457
  • Applications5,325
  • Concepts5
  • Hardware1,802
  • Industry6,143
  • Local Ai4,863
  • Model Releases23,393
  • Research19,837
  • Safety13,180
  • Syntheses17
  • Tools1,671
  • Tutorials3,349

Source
Human
87,042Total entries
1Added by human
87,041Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
61,904 results
29 Jun 2026

Hybrid Fact-Checking that Integrates Knowledge Graphs, Large Language Models, and Search-Based Retrieval Agents Improves Interpretable Claim Verification

Model ReleasesDGX agent

arXiv:2511.03217v2 Announce Type: replace-cross Abstract: Large language models (LLMs) excel in generating fluent utterances but can lack reliable grounding in verified information. At the same time,

HybridCodec: Modeling Discrete and Continuous Representations for Efficient Speech Language Models

ResearchDGX agent

arXiv:2606.27627v1 Announce Type: cross Abstract: Discrete audio representations have become increasingly popular for building multimodal text-audio systems and integrating audio capabilities into Lar

Hyperellipsoid Density Sampling: Exploitative Sequences to Accelerate High-Dimensional Optimization

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases
DGX agent

arXiv:2511.07836v4 Announce Type: replace-cross Abstract: The curse of dimensionality remains a persistent challenge in modern optimization problems. Expanding the search space into higher dimensions

iCost: A Novel Instance-Complexity-Based Cost-Sensitive Learning Framework

Model ReleasesDGX agent

arXiv:2409.13007v3 Announce Type: replace-cross Abstract: Class imbalance poses a significant challenge in classification tasks, often causing standard learning algorithms to become biased toward the

Image-based Geo-localization for Robotics: Are Black-box Vision-Language Models there yet?

Local AiDGX agent

arXiv:2501.16947v2 Announce Type: replace Abstract: The advances in Vision-Language models (VLMs) offer exciting opportunities for robotic applications involving image geo-localization - the problem o

Improving Adversarial Robustness via Activation Amplification and Attenuation

TutorialsDGX agent

arXiv:2606.27784v1 Announce Type: cross Abstract: The existence of adversarial attacks is often attributed to the presence of non-robust features in neural networks. While prior defenses reduce their

Instant Expressive Gaussian Head Avatars at Over 100 FPS

ApplicationsDGX agent

arXiv:2512.16893v2 Announce Type: replace Abstract: Portrait animation has witnessed tremendous quality improvements thanks to recent advances in video diffusion models. However, these 2D methods ofte

Internalizing the Future: A Unified Agentic Training Paradigm for World Model Planning

SafetyDGX agent

arXiv:2606.27483v1 Announce Type: new Abstract: Large language model (LLM) agents have demonstrated strong capability in sequential decision-making, yet they remains fundamentally reactive in long-hor

IWP: Token Pruning as Implicit Weight Pruning in Large Vision Language Models

ResearchDGX agent

arXiv:2604.00757v2 Announce Type: replace-cross Abstract: Large Vision Language Models show impressive performance across image and video understanding tasks, yet their computational cost grows rapidl

JD Oxygen AI Item Center (Oxygen AIIC) V1: An Industrial-Scale LLM/VLM-Centric Solution for Item Understanding, Management, and Applications

ApplicationsDGX agent

arXiv:2606.28070v1 Announce Type: new Abstract: JD.com, one of the world's largest e-commerce platforms, serves over 700 million active users and millions of merchants, with a catalog of tens of billi

Joint Transcription and Decryption of Images of Encrypted Handwritten Documents: A Comparison with the Traditional Pipeline

ApplicationsDGX agent

arXiv:2606.27700v1 Announce Type: cross Abstract: Historical encrypted manuscripts present a challenging problem at the intersection of cryptology, linguistics, paleography, and computer vision. Curre

Just Ask: Curious Code Agents Reveal System Prompts in Frontier LLMs

SafetyDGX agent

arXiv:2601.21233v2 Announce Type: replace Abstract: Autonomous code agents built on large language models are reshaping software and AI development through tool use, long-horizon reasoning, and self-d

KG2Cypher: Data-Centric Pipeline for Building Enterprise Text-to-Cypher Systems

ApplicationsDGX agent

arXiv:2606.27742v1 Announce Type: cross Abstract: Enterprise Knowledge Graphs (KGs) are increasingly used for internal search, analytics, and question answering, but building natural-language interfac

KISS-IMU: Self-supervised Inertial Odometry with Motion-balanced Learning and Uncertainty-aware Inference

ApplicationsDGX agent

arXiv:2603.06205v2 Announce Type: replace Abstract: Inertial measurement units (IMUs), which provide high-frequency linear acceleration and angular velocity measurements, serve as fundamental sensing

Ko-WideSearch: A Korean Breadth-Search Benchmark for Exhaustive Set Enumeration by Web Agents

Model ReleasesDGX agent

arXiv:2606.27595v1 Announce Type: new Abstract: Web-agent benchmarks overwhelmingly measure depth -- pinning one obscure answer behind a chain of constraints -- while breadth, exhaustively enumerating

Large Language Model Teaches Visual Students: Cross-Modality Transfer of Fine-Grained Conceptual Knowledge

ResearchDGX agent

arXiv:2606.27527v1 Announce Type: cross Abstract: Large Language Models (LLMs) possess broad conceptual knowledge acquired through large-scale text pretraining, yet their potential to supervise models

Latent Visual Diffusion Reasoning with Monte Carlo Tree Search

TutorialsDGX agent

arXiv:2606.27988v1 Announce Type: new Abstract: Analyzing fine-grained skill activities (e.g., sports, surgery) requires not only recognizing visual patterns but also performing step-by-step visual re

Layerwise Progressive Freezing: A Training Scaffold for Depth-Scalable Binary Networks

ResearchDGX agent

arXiv:2606.27759v1 Announce Type: new Abstract: Training binary neural networks (BNNs) from scratch is dominated by the straight-through estimator (STE), whose forward/backward mismatch produces sever

Learn Structure, Adapt on the Fly: Multi-Scale Residual Learning and Online Adaptation for Aerial Manipulators

AgentsDGX agent

arXiv:2603.11638v2 Announce Type: replace Abstract: Autonomous Aerial Manipulators (AAMs) are inherently coupled, nonlinear systems that exhibit nonstationary and multiscale residual dynamics, particu

Learning 1-Bit LiDAR-based Localization with Auxiliary Objective

AgentsDGX agent

arXiv:2606.27729v1 Announce Type: new Abstract: 6-DoF LiDAR-based localization is a fundamental capability for autonomous systems operating in large-scale outdoor environments. Many deep-learning-base

Learning Complementary Action Modeling from Automotive Maintenance Instructions

ResearchDGX agent

arXiv:2606.27808v1 Announce Type: new Abstract: A minute lexical variation can reverse the procedural meaning of an instruction even when the rest of the sentence remains unchanged. In automotive main

Learning from Annotation Uncertainty: Entropy-Aware Curriculum for Speech Emotion Recognition

SafetyDGX agent

arXiv:2606.27536v1 Announce Type: cross Abstract: Speech emotion recognition (SER) often relies on hard consensus labels that collapse annotator disagreement. We study distribution-based supervision f

Learning in Markovian bandits with non-observable states and constrained decision epochs

Model ReleasesDGX agent

arXiv:2606.27448v1 Announce Type: new Abstract: This paper studies the problem of regret minimization in Markovian bandits with non-observable states and possibly constrained decision epochs. The focu

Learning Peer Influence Probabilities with Linear Contextual Bandits

Model ReleasesDGX agent

arXiv:2510.19119v2 Announce Type: replace Abstract: In networked environments, it is common for users to share recommendations about content, products, services, and possible courses of action. Whethe

Learning Stable In-Grasp Manipulation in a Non-Dropping Action Space

ResearchDGX agent

arXiv:2606.28196v1 Announce Type: new Abstract: Traditionally, dexterous manipulation controllers are designed using analytic models constrained by strong assumptions about the hand and the objects be

Learning Stochastic Bridges for Video Object Removal via Video-to-Video Translation

ResearchDGX agent

arXiv:2601.12066v4 Announce Type: replace Abstract: Existing video object removal methods predominantly rely on diffusion models following a noise-to-data paradigm, where generation starts from uninfo

Learning to Evict from Key-Value Cache

Model ReleasesDGX agent

arXiv:2602.10238v2 Announce Type: replace Abstract: The growing size of Large Language Models (LLMs) makes efficient inference challenging, primarily due to the memory demands of the autoregressive Ke

Learning to Reason with Curriculum II: Compositional Generalization

ResearchDGX agent

arXiv:2606.27721v1 Announce Type: new Abstract: Compositional generalization, the ability to solve complex problems by combining solutions to simpler sub-problems, is a fundamental capability of both

Learning to Refine Hidden States for Reliable LLM Reasoning

SafetyDGX agent

arXiv:2606.17524v2 Announce Type: replace Abstract: Large language models show strong reasoning ability, but their internal reasoning process can remain unstable in complex multi-step settings, where

Learning to Throw: Agile and Accurate Cable-Suspended Payload Delivery with a Quadrotor

SafetyDGX agent

arXiv:2606.27603v1 Announce Type: new Abstract: Quadrotors offer the agility needed to rapidly transport suspended payloads during time-critical applications, including search-and-rescue and medical d

Learning Topology-Aware Representations via Test-Time Adaptation for Anomaly Segmentation

TutorialsDGX agent

arXiv:2606.28268v1 Announce Type: cross Abstract: Test-time adaptation (TTA) has emerged as a promising paradigm for mitigating distribution shifts in deep models. However, existing TTA approaches for

Let Language Constrain Geometry: Vision-Language Models as Semantic and Spatial Critics for 3D Generation

SafetyDGX agent

arXiv:2511.14271v2 Announce Type: replace Abstract: Text-to-3D generation has advanced rapidly, yet state-of-the-art models, encompassing both optimization-based and feed-forward architectures, still

LieSolver: PDE-Constrained Learning for IBVPs via Lie Symmetries

ResearchDGX agent

arXiv:2510.25731v2 Announce Type: replace-cross Abstract: Initial-boundary value problems (IBVPs) provide the essential framework for modelling a wide range of phenomena in physics and engineering. We

Lifted Causal Inference

ResearchDGX agent

arXiv:2606.28024v1 Announce Type: new Abstract: Lifted inference exploits indistinguishabilities in probabilistic graphical models by using a representative for indistinguishable objects, thereby spee

LLawCo: Learning Laws of Cooperation for Modeling Embodied Multi-Agent Behavior

Model ReleasesDGX agent

arXiv:2606.28182v1 Announce Type: cross Abstract: Embodied agents operating in decentralized and partially observable environments have attracted growing attention in recent years. However, existing l

LocalNav: Distilling Frontier VLMs and Embodied RL for On-Device Object Goal Navigation

Model ReleasesDGX agent

arXiv:2606.27871v1 Announce Type: new Abstract: Vision Language Models (VLMs) have emerged in the robotic domain as a powerful tool that enables environmental perception with language context, serving

Long-Term Prediction of Local and Global Human Motion with Occlusion Recovery

Local AiDGX agent

arXiv:2606.27900v1 Announce Type: new Abstract: Human motion describes the three-dimensional full-body movement of a person. Anticipating such motion holds significant relevance across a wide range of

Lost at the End: Primacy Bias in Multimodal Retrieval-Augmented Question Answering

Model ReleasesDGX agent

arXiv:2606.16494v2 Announce Type: replace-cross Abstract: Knowledge-based visual question answering (KB-VQA) lets vision-language systems answer questions that exceed their parametric knowledge by con

Low-Agreeableness Persona Conditioning for Safe LLM Fine-Tuning

SafetyDGX agent

arXiv:2606.27709v1 Announce Type: cross Abstract: Recent work has shown that fine-tuning large language models (LLMs) for social warmth degrades factual reliability and increases sycophancy. We invest

LXD-SLAM: LiDAR+X Dense SLAM with sum_{i=0}^{5}C_5^i Configurable Sensor Combinations

AgentsDGX agent

arXiv:2606.27811v1 Announce Type: new Abstract: Simultaneous Localization and Mapping (SLAM) is essential for autonomous systems, yet achieving reliable, globally consistent pose estimation and dense

Masked Language Flow Models

ResearchDGX agent

arXiv:2606.27617v1 Announce Type: new Abstract: Masked Diffusion Models (MDMs) promise fast, parallel language generation, but their reverse transition factorises across token positions -- an approxim

MASS: Motion-Aligned Selective Scan for Refinement in Flow-Based Video Frame Interpolation

TutorialsDGX agent

arXiv:2606.27718v1 Announce Type: new Abstract: Video frame interpolation (VFI) remains a challenging task, particularly when dealing with large, non-linear motions and complex occlusions. While flow-

Measuring the Redundancy of Decoder Layers in SpeechLLMs

ResearchDGX agent

arXiv:2603.05121v2 Announce Type: replace-cross Abstract: Speech Large Language Models route speech encoder representations into an LLM decoder that typically accounts for over 90% of total parameters

Mechanism-Driven Monitors for Preemptive Detection of LLM Training Instability

ResearchDGX agent

arXiv:2606.28116v1 Announce Type: new Abstract: Frontier large language model training consumes massive accelerator fleets and long wall-clock computation, making stability failures costly when they o

MeDUET: Disentangled Unified Pretraining for 3D Medical Image Synthesis and Analysis

TutorialsDGX agent

arXiv:2602.17901v3 Announce Type: replace-cross Abstract: Self-supervised learning (SSL) and diffusion models have respectively advanced representation learning and generative modeling for high-dimens

MemoBench: Benchmarking World Modeling in Dynamically Changing Environments

Model ReleasesDGX agent

arXiv:2606.27537v1 Announce Type: new Abstract: Video generation models aspire to simulate dynamic environments, and several benchmarks now evaluate memory consistency across frames. However, most ass

MER-R1: Multimodal Emotion Reasoning via Slow-Fast Thinking Synergy

ResearchDGX agent

arXiv:2606.27652v1 Announce Type: new Abstract: We find that explicit reasoning does not necessarily translate into better multimodal emotion recognition (MER) accuracy, even though it makes predictio

MetaBreak: Jailbreaking Online LLM Services via Special Token Manipulation

SafetyDGX agent

arXiv:2510.10271v2 Announce Type: replace-cross Abstract: Unlike regular tokens derived from existing text corpora, special tokens are artificially created to annotate structured conversations during

Mind the Gap: Quantifying the Domain Gap in Cross-Sensor Diffusion Super-Resolution

ResearchDGX agent

arXiv:2606.28039v1 Announce Type: cross Abstract: Demand for high-resolution satellite imagery has increased interest in super-resolution (SR) to bridge the spatial resolution gap between freely avail

MindFlow: Harmonizing Cognitive Semantics and Acoustic Dynamics for Facial Animation Generation in Dyadic Conversations

ResearchDGX agent

arXiv:2606.27779v1 Announce Type: new Abstract: Generating lifelike facial animation for dyadic conversations requires reconciling high-level cognitive intent with precise low-level motor reflexes, ye

Mitigating LLM-based p-Hacking by Preregistering for the Next LLM

Model ReleasesDGX agent

arXiv:2606.27687v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used to generate, classify, and annotate data whose outputs feed downstream hypothesis tests. However, L

Mitigating Position Bias in Transformers via Layer-Specific Positional Embedding Scaling

SafetyDGX agent

arXiv:2606.27705v1 Announce Type: new Abstract: Large Language Models (LLMs) still struggle with the ``lost-in-the-middle'' problem, where critical information located in the middle of long-context in

MixTTA: Low-Rank Cross-Channel Mixing for Reliable Test-Time Adaptation

ResearchDGX agent

arXiv:2606.28142v1 Announce Type: new Abstract: Test-Time Adaptation (TTA) methods commonly update the affine parameters of normalization layers to adapt deployed models under distribution shifts. How

MLVC: Multi-platform Learned Video Codec for Real-World Deployment

Model ReleasesDGX agent

arXiv:2606.28027v1 Announce Type: cross Abstract: Neural video codecs have surpassed classical codecs in coding efficiency but remain impractical for deployment due to cross-platform incompatibility a

MobileManiBench: Simplifying Model Verification for Mobile Manipulation

Model ReleasesDGX agent

arXiv:2602.05233v2 Announce Type: replace Abstract: Vision-language-action models have advanced robotic manipulation but remain constrained by reliance on the large, teleoperation-collected datasets d

ModaFlow: Modality-Aware Flow Matching for High-Fidelity Virtual Try-On

SafetyDGX agent

arXiv:2606.27773v1 Announce Type: new Abstract: Image-based virtual try-on has emerged as a compelling task in e-commerce and augmented reality, yet existing methods struggle to simultaneously preserv

Monocular Avatar Reconstruction via Cascaded Diffusion Priors and UV-Space Differentiable Shading

Model ReleasesDGX agent

arXiv:2606.28144v1 Announce Type: new Abstract: Reconstructing high-fidelity, relightable 3D avatars from a single in-the-wild image is a challenging ill-posed problem, primarily hindered by the scarc

Monte Carlo with kernel-based Gibbs measures: Guarantees for probabilistic herding

Model ReleasesDGX agent

arXiv:2402.11736v3 Announce Type: replace Abstract: Kernel herding belongs to a family of deterministic quadratures that seek to minimize the maximum mean discrepancy (MMD), that is, the worst-case in

Mosaic: A Benchmark Suite for Differentiable Physics Solvers

Model ReleasesDGX agent

arXiv:2606.27895v1 Announce Type: cross Abstract: Differentiable partial differential equation (PDE) solvers underpin solver-in-the-loop ML training, gradient-based optimal control, and inverse proble

MPFlow: Multi-modal Posterior-Guided Flow Matching for Zero-Shot MRI Reconstruction

SafetyDGX agent

arXiv:2603.03710v3 Announce Type: replace-cross Abstract: Zero-shot MRI reconstruction relies on generative priors, but single-modality unconditional priors produce hallucinations under severe ill-pos

← Previous
1…340341342343344…1032
Next →