AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,403
  • Agents7,557
  • Applications5,411
  • Concepts5
  • Hardware1,836
  • Industry6,170
  • Local Ai4,931
  • Model Releases23,900
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,403
  • Agents7,557
  • Applications5,411
  • Concepts5
  • Hardware1,836
  • Industry6,170
  • Local Ai4,931
  • Model Releases23,900
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
Human
88,403Total entries
1Added by human
88,402Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
62,897 results
3 Jun 2026

ASymPO: Asymmetric-Scale Policy Optimization for Asynchronous LLM Post-Training Without Behavior Information

SafetyDGX agent

arXiv:2606.03070v1 Announce Type: cross Abstract: Asynchronous reinforcement learning can improve language-model post-training throughput by decoupling response generation from policy optimization, bu

ATLAS: A Large-Scale Evaluation Benchmark for Adversarial LiDAR Perception

Model ReleasesDGX agent

arXiv:2606.02924v1 Announce Type: new Abstract: Autonomous driving perception is typically evaluated on clean benchmark data, yet real-world deployment requires robustness to rare, structured, and pot

Attend to Anything: Foundation Model for Unified Human Attention Modeling

ApplicationsDGX agent

arXiv:2606.03540v1 Announce Type: new Abstract: Existing human attention (saliency) modeling methods persist as highly fragmented across modalities, scenes, and task formulations. Consequently, even w

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Attention Calibration for Position-Fair Dense Information Retrieval

SafetyDGX agent

arXiv:2606.02737v1 Announce Type: cross Abstract: Dense retrieval models exhibit positional bias: retrieval effectiveness degrades when relevant information appears later in a passage (Zeng et al., 20

Attention, May I Have Your Decision? Localizing Generative Choices in Diffusion Models

Local AiDGX agent

arXiv:2604.06052v2 Announce Type: replace Abstract: Text-to-image diffusion models exhibit remarkable generative capabilities, yet their internal operations remain opaque, particularly when handling p

Attribution via Distributional Paths for Information Revelation

ResearchDGX agent

arXiv:2606.03885v1 Announce Type: new Abstract: Feature attribution methods explain predictions by assigning importance scores to input features. Path-based methods such as Integrated Gradients are es

Auditable Climate Risk Intelligence from Fragmented ESG Data: Deterministic Orchestration and Imbalance-Aware Learning for Scope 1-3 Validation

Model ReleasesDGX agent

arXiv:2606.02604v1 Announce Type: cross Abstract: ESG and climate risk data remain fragmented across heterogeneous Scope 1, Scope 2, and Scope 3 reporting environments, while conventional validation p

AUDITFLOW: Executable Symbolic Environments for Structured Financial Reporting Verification

Model ReleasesDGX agent

arXiv:2606.03031v1 Announce Type: new Abstract: Structured financial audit verification is difficult for language-model agents because correctness depends on structured evidence rather than text alone

Auditing Engagement Incentives in the Kidfluencer Ecosystem: A Multimodal Weak Supervision Approach

Model ReleasesDGX agent

arXiv:2606.03173v1 Announce Type: cross Abstract: The rise of `kidfluencers' on YouTube has raised ethical concerns about child digital labor and exploitation. While emerging legislation attempts to r

AugMask: Training Diffusion Models on Incomplete Tabular Data via Stochastic Augmentation and Masking

ApplicationsDGX agent

arXiv:2606.03347v1 Announce Type: cross Abstract: Score-based diffusion models have emerged as prominent deep generative models; however, their application to tabular data remains challenging because

AUGUSTE: Online-Learning dApp for Predictive URLLC Scheduling

AgentsDGX agent

arXiv:2606.03664v1 Announce Type: cross Abstract: Ultra Reliable and Low Latency Communications (URLLC) was one of the main motivations behind 5G, with 3GPP advertising 1-10 ms latency targets for app

AURA: Action-Gated Memory for Robot Policies at Constant VRAM

Model ReleasesDGX agent

arXiv:2606.02775v1 Announce Type: new Abstract: The KV-cache is the right memory for datacenters but the wrong memory for robots. Datacenter inference batches many short requests and resets them, amor

Automata-Conditioned Cooperative Multi-Agent Reinforcement Learning

AgentsDGX agent

arXiv:2511.02304v2 Announce Type: replace-cross Abstract: We study learning multi-task, multi-agent policies for cooperative, temporal objectives, under centralized training, decentralized execution.

Automated Report-Derived Oncology VQA Benchmark for Evaluating Vision-Language Models on 3D Medical Imaging

Model ReleasesDGX agent

arXiv:2606.02809v1 Announce Type: new Abstract: Evaluating vision-language models (VLMs) on medical images requires benchmarks that are clinically grounded, scalable, and controlled for evaluation con

Autonomous Navigation System for Library Service Robot Based on Unitree Go2 Edu

AgentsDGX agent

arXiv:2606.03340v1 Announce Type: new Abstract: Libraries require autonomous robots to move quietly through narrow aisles while remaining safe around readers, chairs, bags, and carts. This paper prese

AutoTail-BSFGM: Class-Balance-Aware Fine-Tuning for Chinese Scholarly Text Classification

ResearchDGX agent

arXiv:2606.03576v1 Announce Type: new Abstract: Scholarly text classification supports literature organization, subject indexing, and research intelligence, but Chinese scholarly corpora often contain

AvatarMix: Identity-Preserving Cross-Avatar Composition for Outfit Personalization

ResearchDGX agent

arXiv:2606.03506v1 Announce Type: new Abstract: Existing 3D avatar outfit transfer methods face distinct challenges: approaches that lift 2D edits to 3D often suffer from outfit or identity quality de

AVTrack: Audio-Visual Tracking in Human-centric Complex Scenes

Model ReleasesDGX agent

arXiv:2606.02724v1 Announce Type: cross Abstract: Audio-visual speaker tracking aims to localize and track active speakers by leveraging auditory and visual cues, enabling fine-grained, human-centric

BA-T: An Iterative Transformer for Two-View Bundle Adjustment

ResearchDGX agent

arXiv:2606.03287v1 Announce Type: new Abstract: Feed-forward models for 3D reconstruction have achieved strong performance using deep cross-view attention to exchange information across images. Howeve

Backdoor Unlearning Generalization: A Path Toward the Removal of Unknown Triggers in LLMs

SafetyDGX agent

arXiv:2606.03785v1 Announce Type: new Abstract: Backdoor attacks in Large Language Models (LLMs) are a growing security concern, where models can generate adversary-chosen content. Existing defenses t

BAHSD: Bridging the Long-tail Gap via Adaptive Distillation in Black-box Sequential Recommendation

ResearchDGX agent

arXiv:2606.03091v1 Announce Type: cross Abstract: Sequential recommendation systems are widely adopted but often deployed as black-box APIs, which has driven recent interest in model extraction to rep

Balancing Symmetry and Efficiency in Graph Flow Matching

ResearchDGX agent

arXiv:2602.18084v2 Announce Type: replace Abstract: Equivariance is central to graph generative models, as it ensures the model respects the permutation symmetry of graphs. However, strict equivarianc

BaltiVoice: A Speech Corpus and Fine-tuned Whisper ASR System for the Balti Language

ResearchDGX agent

arXiv:2606.03504v1 Announce Type: cross Abstract: We present BaltiVoice, a 16.8-hour read-speech corpus for Balti (ISO 639-3: bft), a Tibetic language spoken in Gilgit-Baltistan, Pakistan, with no pri

Bayesian Tensor Decomposition with Diffusion Model Prior

SafetyDGX agent

arXiv:2606.03212v1 Announce Type: new Abstract: Low-rank tensor decomposition (TD) is usually effective on clean, fully observed data, but it often degrades under severe missingness or noise. Low-rank

BEAST3D: Animal behavioral analysis and neural encoding from multi-view video via Gaussian splatting

ResearchDGX agent

arXiv:2606.02937v1 Announce Type: cross Abstract: Multi-view video recordings are increasingly used to capture the 3D movements of animals in experimental settings, yet extracting rich 3D representati

Before Fusion, Ask What to Keep: Contextual Calibration of Multimodal Signals

Local AiDGX agent

arXiv:2606.02679v1 Announce Type: new Abstract: Multimodal systems often benefit from combining information across language, sound, and visual streams, but this benefit is not guaranteed. A modality t

BehaviorBench: Modeling Real-World User Decisions from Behavioral Traces

Model ReleasesDGX agent

arXiv:2606.02798v1 Announce Type: new Abstract: Many decision-support settings require systems that adapt to individual users, but evaluation data for this problem remain limited. Existing benchmarks

Benchmarking Speech-to-Speech Translation Models

ApplicationsDGX agent

arXiv:2606.03241v1 Announce Type: new Abstract: Speech-to-speech translation (S2ST) has advanced rapidly, but offline evaluation lacks a unified protocol: studies report non-overlapping metric subsets

Benchmarking Visual State Tracking in Multimodal Video Understanding

Model ReleasesDGX agent

arXiv:2606.03920v1 Announce Type: new Abstract: Understanding a video requires more than recognizing isolated moments, as humans continuously track entities, states, and events over time. This capacit

Best of Both Worlds: Multimodal Reasoning and Generation via Unified Discrete Flow Matching

SafetyDGX agent

arXiv:2602.12221v2 Announce Type: replace Abstract: We propose UniDFlow, a unified discrete flow-matching framework for multimodal understanding, generation, and editing. It decouples understanding an

BEV-ODOM2: Enhanced BEV-based Monocular Visual Odometry with PV-BEV Fusion and Dense Flow Supervision for Ground Robots

Model ReleasesDGX agent

arXiv:2509.14636v2 Announce Type: replace Abstract: Scale-consistent ego-motion estimation is fundamental for autonomous ground robots. Bird's-Eye-View (BEV) representation naturally addresses the sca

Beyond Compression: Quantifying Spectral Accessibility in Vision Representations

ResearchDGX agent

arXiv:2606.03795v1 Announce Type: new Abstract: Vision-language models map visual features into a shared embedding space through learned projection layers, yet it remains unclear how these transformat

Beyond Encoder Accumulation: Measuring Encoder Roles in Multi-Encoder VLMs

Model ReleasesDGX agent

arXiv:2606.03879v1 Announce Type: cross Abstract: As foundation models scale toward fusing more heterogeneous visual streams, understanding how diverse encoders interact under joint training becomes a

Beyond False Stability: High-Noise Drift Gating for Test-Time Adversarial Defenses in Vision-Language Models

ResearchDGX agent

arXiv:2606.03730v1 Announce Type: new Abstract: Vision-language models (VLMs) such as CLIP show strong zero-shot generalization but remain highly vulnerable to adversarial attacks. Adversarial trainin

Beyond Gradient Descent: Adam for Analog Ising Machines

ResearchDGX agent

arXiv:2606.03917v1 Announce Type: cross Abstract: As Moore's law reaches its limits, Ising machines offer a promising alternative computing approach for difficult optimization problems. However, many

Beyond Ideal Instruction: A Comprehensive Framework for Evaluating LLMs in Realistic Interactions

Model ReleasesDGX agent

arXiv:2606.03318v1 Announce Type: new Abstract: Despite great advances in tool-use capabilities of large language models (LLMs), existing evaluation benchmarks struggle to fully align with real-world

Beyond Semantics: Modeling Factual and Affective Perceptual Experiences from Vision-Language Data

ResearchDGX agent

arXiv:2606.03345v1 Announce Type: cross Abstract: We present P-Topics (Perception Topics) modeling, a novel problem for understanding how images are perceived affectively and across cultures. The goal

Beyond Single Solution: Multi-Hypothesis Collaborative Deep Unfolding Network for Image Compressive Sensing

ResearchDGX agent

arXiv:2606.03666v1 Announce Type: new Abstract: Recent deep unfolding networks (DUNs) have advanced Compressive Sensing (CS) by effectively integrating iterative optimization with deep learning archit

Beyond the Literal: Decomposing Pragmatic Intent in Multimodal Meme Understanding

ResearchDGX agent

arXiv:2606.03604v1 Announce Type: new Abstract: When asked what a meme or sarcastic post means, Large Vision Language Models (LVLMs) tend to describe what the image shows rather than what the author i

Beyond 'To whom it may concern': Tailoring Machine Translation to Audience and Intent

ResearchDGX agent

arXiv:2606.03259v1 Announce Type: new Abstract: Translation quality depends on purpose: the same source text demands different translations depending on audience, tone, and communicative intent. Yet M

BigFinanceBench: A Workflow-Grounded Benchmark for Financial-Research Agents

Model ReleasesDGX agent

arXiv:2606.03829v1 Announce Type: new Abstract: Financial-research answers are decision-relevant only when another analyst can audit how they were produced: which source was chosen, which period and a

Binary Road Surface Classification Using Machine Learning on Production Vehicle Signals During Cruising

ApplicationsDGX agent

arXiv:2606.02762v1 Announce Type: new Abstract: Knowledge of real-time road slipperiness, or even better, a refined estimate of peak grip potential, is a critical input for vehicle warning and interve

Bionic Human-Motion Style Transfer for Physically Executable Whole-Body Control of Humanoid Robots

SafetyDGX agent

arXiv:2606.03536v1 Announce Type: new Abstract: Expressive whole-body motion is important for humanoid robots operating in human environments, where robots are expected to move stably while presenting

Black-box, Adaptive, Efficient, Transferable, Harmful, Applicable... Attacks Are All You Need to Break LLMs

SafetyDGX agent

arXiv:2606.03647v1 Announce Type: cross Abstract: Accurately evaluating adversarial robustness is a longstanding challenge. A flawed attack design can inflate robustness estimates, making deployment r

Bootstrap Your Generator: Unpaired Visual Editing with Flow Matching

ResearchDGX agent

arXiv:2606.03911v1 Announce Type: new Abstract: Modern generative models possess a deep understanding of visual content, yet training them for image editing typically requires massive datasets of pair

BotDirector: Robot Storytelling Across the Symmetrical Reality with Multi-modal Interactions

AgentsDGX agent

arXiv:2606.03223v1 Announce Type: cross Abstract: Robot storytelling offers a unique blend of technological innovation and creative expression that engages children in unprecedented ways. However, the

Breaking the Self-Confirming Loop: Diagnosing and Mitigating Systemic Reward Bias in Self-Rewarding RL

SafetyDGX agent

arXiv:2510.08977v2 Announce Type: replace-cross Abstract: Reinforcement learning with verifiable rewards (RLVR) efficiently scales the reasoning ability of large language models (LLMs) but is bottlene

Bregman meets Levy: Stochastic mirror descent with heavy-tailed noise in continuous and discrete time

ResearchDGX agent

arXiv:2606.03769v1 Announce Type: cross Abstract: We study the robustness of stochastic mirror descent (SMD) under heavy-tailed noise, focusing on whether the method retains its convergence guarantees

Bridging Auxiliary Constraints to Resolve Instruction Following in Large Reasoning Models

ResearchDGX agent

arXiv:2606.03624v1 Announce Type: new Abstract: Large Reasoning Models (LRMs) have demonstrated impressive capabilities in many tasks, yet they struggle with reliably following multiple instructions,

Bridging Predictive Uncertainty and Safe Action: Sample-Conditioned Differentiable Planning for Autonomous Driving

SafetyDGX agent

arXiv:2606.03296v1 Announce Type: new Abstract: Complex, dynamic, and interactive driving environments pose significant challenges for autonomous driving, primarily due to the pervasive uncertainty of

Brief Announcement: Generative Markov Model for Distributed Computing Systems

SafetyDGX agent

arXiv:2606.03061v1 Announce Type: cross Abstract: Emerging distributed computing paradigms, such as the computing continuum, are inherently heterogeneous, stochastic, and complex. Efficiently and effe

Building Better Activation Oracles

SafetyDGX agent

arXiv:2606.02609v1 Announce Type: cross Abstract: Activation Oracles (AOs) are promising methods for interpreting residual stream activations. However, current AOs face important issues, such as hallu

Building Reliable Long-Form Generation via Hallucination Rejection Sampling

ResearchDGX agent

arXiv:2606.03628v1 Announce Type: cross Abstract: Large language models (LLMs) have achieved remarkable progress in open-ended text generation, yet they remain prone to hallucinating incorrect or unsu

Building Trust in Black-box Optimization: A Comprehensive Framework for Explainability

ApplicationsDGX agent

arXiv:2410.14573v2 Announce Type: replace-cross Abstract: Optimizing costly black-box functions within a constrained evaluation budget presents significant challenges in many real-world applications.

Buzz, Choose, Forget: A Meta-Bandit Framework for Bee-Like Decision Making

ResearchDGX agent

arXiv:2510.16462v3 Announce Type: replace Abstract: This work introduces MAYA, a sequential imitation learning model based on multi-armed bandits, designed to reproduce and predict individual bees' de

BYORn: Bootstrap Your Own Responses to Defend Large Vision-Language Models Against Backdoor Attacks

ResearchDGX agent

arXiv:2606.02947v1 Announce Type: cross Abstract: Supervised fine-tuning is the predominant approach for adapting autoregressive vision-language models to downstream tasks. Recent work has shown that

CAD-to-CT Registration of Cylindrical Objects via Ellipse-Based Axis Estimation

ResearchDGX agent

arXiv:2606.02935v1 Announce Type: new Abstract: Accurate registration of CAD models to CT scans is essential for establishing ground truth geometry in volumetric imaging. Obtaining reliable object mas

Calibrating Urban Traffic Simulation from Sparse Road Observations via Genetic Optimization

ApplicationsDGX agent

arXiv:2606.03823v1 Announce Type: new Abstract: Urban traffic simulation is a critical tool for infrastructure planning, including the placement of electric vehicle charging stations. However, realist

Calibration Data Trade-offs Across Capability Dimensions: Why Multi-Source Mixing Matters for High-Sparsity LLM Pruning

Model ReleasesDGX agent

arXiv:2606.03328v1 Announce Type: cross Abstract: Post-training pruning compresses large language models to high sparsity using a small unlabelled calibration set, and recent work has concluded that t

Can Factual Opinions Be Edited (Manipulated) in Large Language Models?

Model ReleasesDGX agent

arXiv:2606.03096v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly integrated into various domains, making knowledge editing techniques crucial yet potentially hazardous. Cu

← Previous
1…491492493494495…1049
Next →