AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,271
  • Agents7,543
  • Applications5,408
  • Concepts5
  • Hardware1,828
  • Industry6,162
  • Local Ai4,927
  • Model Releases23,818
  • Research20,122
  • Safety13,367
  • Syntheses17
  • Tools1,674
  • Tutorials3,400

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,271
  • Agents7,543
  • Applications5,408
  • Concepts5
  • Hardware1,828
  • Industry6,162
  • Local Ai4,927
  • Model Releases23,818
  • Research20,122
  • Safety13,367
  • Syntheses17
  • Tools1,674
  • Tutorials3,400

Source
Human
88,271Total entries
1Added by human
88,270Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
62,897 results
11 Jun 2026

Towards a Bridge Layer Between Bibliographic and Formalized Mathematical Knowledge

SafetyDGX agent

arXiv:2606.11430v1 Announce Type: cross Abstract: Mathematical knowledge is split between bibliographic databases (e.g., MathSciNet, zbMATH Open) and formal proof libraries (e.g., Lean mathlib), preve

Towards Conditional Feature Alignment for Cross-Domain Counting

SafetyDGX agent

arXiv:2506.17137v3 Announce Type: replace Abstract: Object counting models often degrade under cross-domain deployment because density composition varies across domains and is itself task-relevant. St

Towards Data-free and Training-free Compression for Speech Foundation Models Using Parameter Clustering

Model ReleasesDGX agent

arXiv:2606.11836v1 Announce Type: cross Abstract: This paper presents a novel data-free and training-free compression approach for speech foundation models using channelwise clustering via k-means. Mo

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Towards Deep Learning Surrogate for the Forward Problem in Electrocardiology: A Scalable Alternative to Physics-Based Models

ResearchDGX agent

arXiv:2512.13765v2 Announce Type: replace-cross Abstract: The forward problem in electrocardiology, computing body surface potentials from cardiac electrical activity, is traditionally solved using ph

Towards Fully Automated Exam Grading: Fairness-Aware Recognition of Handwritten Answers with Foundation Models

Model ReleasesDGX agent

arXiv:2606.11477v1 Announce Type: cross Abstract: Correcting handwritten exams by hand is time-consuming and error-prone, particularly for large cohorts, while fully digital exams tend to force a dida

Towards Responsibly Non-Compliant Machines

AgentsDGX agent

arXiv:2606.12147v1 Announce Type: new Abstract: We consider the problem of engineering autonomous intelligent agents that are capable to responsibly not comply with user requests. We argue that machin

Traceable Virtual Sea Trials in the Marine Robotics Unity Simulator for Manoeuvring Assessment of Unmanned Surface Vehicles

SafetyDGX agent

arXiv:2606.12349v1 Announce Type: new Abstract: Accurate identification of hydrodynamic derivatives is essential for control and navigation of Unmanned Surface Vehicles (USVs), but high-fidelity manoe

Traits Run Deeper: Trait-Specific Asymmetric Fusion for Personality Assessment

SafetyDGX agent

arXiv:2606.11269v1 Announce Type: new Abstract: Personality assessment aims to infer stable personality traits from dynamic behaviors across language, voice, and facial cues. Since different personali

Tree-Structured Orthonormal Decomposition of the Aitchison Simplex

Local AiDGX agent

arXiv:2606.11646v1 Announce Type: new Abstract: Compositional data -- vectors encoding relative proportions -- arise across scientific domains, including ecology, geochemistry, and genomics. The featu

TreeSeeker: Tree-Structured Trial, Error, and Return in Deep Search

AgentsDGX agent

arXiv:2606.11662v1 Announce Type: new Abstract: Deep search requires agents to answer complex questions through multi-step web search, browsing, evidence comparison, and synthesis. A central challenge

TRON: Tracing Rays to Orchestrate a Neural Renderer for 3D Gaussian Reconstructions

ApplicationsDGX agent

arXiv:2606.11314v1 Announce Type: new Abstract: We introduce TRON, a rendering framework that combines 3D Gaussian ray tracing with neural rendering to enable realistic and controllable rendering of r

UGV-Conditioned Multi-UAV Informative Planning on a Shared Exposure Belief

SafetyDGX agent

arXiv:2606.12306v1 Announce Type: new Abstract: Safe ground navigation in large, threat-augmented environments requires aerial support that actively reduces the risks that a ground vehicle faces along

Understanding Cross-Sensor Feature Variations for Generalizable 3D Perception

ResearchDGX agent

arXiv:2606.11573v1 Announce Type: new Abstract: Radar-camera BEV perception often suffers from degraded performance when evaluated across datasets, as changes in driving scenes, sensor configurations,

Unifying Learning Dynamics and Generalization in Transformers Scaling Law

ApplicationsDGX agent

arXiv:2512.22088v3 Announce Type: replace-cross Abstract: The scaling law, a cornerstone of Large Language Model (LLM) development, predicts improvements in model performance with increasing computati

UniIntervene: Agentic Intervention for Efficient Real-World Reinforcement Learning

SafetyDGX agent

arXiv:2606.12372v1 Announce Type: cross Abstract: Human-in-the-loop reinforcement learning (HiL-RL) has emerged as an effective paradigm for real-world robotic manipulation, enabling online policy imp

UniReason-Med: A Shared Grounded Reasoning Interface for 2D-to-3D Transfer in Medical VQA

SafetyDGX agent

arXiv:2606.11740v1 Announce Type: cross Abstract: We study whether grounded reasoning supervision from abundant 2D medical images can improve 3D medical VQA when both input types are aligned through a

Unstable Features, Reproducible Subspaces: Understanding Seed Dependence in Sparse Autoencoders

ResearchDGX agent

arXiv:2606.12138v1 Announce Type: cross Abstract: Sparse autoencoders (SAEs) are widely used to interpret neural network representations, but their utility depends on whether the learned features are

UR-BERT: Scaling Text Encoders for Massively Multilingual TTS Through Universal Romanization and Speech Token Prediction

SafetyDGX agent

arXiv:2606.11681v1 Announce Type: new Abstract: We propose UR-BERT, a Romanized transcription-based text-to-speech (TTS) encoder for massively multilingual TTS systems. Conventional grapheme-to-phonem

Urban Heat MiniCubes: An AI-Ready dataset for urban heat research

SafetyDGX agent

arXiv:2606.11534v1 Announce Type: cross Abstract: Urban heat is amplified by impermeable surfaces and heterogeneous built environments, yet street-level variability remains difficult to quantify becau

Using Explainability as a Training-Time Reliability Signal for Efficient ECG Classification

Model ReleasesDGX agent

arXiv:2606.12252v1 Announce Type: cross Abstract: Training deep neural networks for clinical time-series analysis is computationally demanding, yet many healthcare settings lack the resources required

uva-irlab-conv at SemEval-2026 Task 8: Multi-Turn RAG with Learned Sparse Retrieval and Listwise Reranking

ApplicationsDGX agent

arXiv:2606.11945v1 Announce Type: new Abstract: This report describes our participation in SemEval-2026 Task 8 on multi-turn retrieval and question answering. The task evaluates conversational systems

Vector Quantized Latent Concepts: A Scalable Alternative to Clustering-Based Concept Discovery

ResearchDGX agent

arXiv:2602.02726v2 Announce Type: replace-cross Abstract: Large language models (LLMs) encode rich semantic information in their hidden states, yet it remains difficult to understand what information

Verifiable Environments Are LEGO Bricks: Recursive Composition for Reasoning Generalization

Model ReleasesDGX agent

arXiv:2606.12373v1 Announce Type: new Abstract: Reinforcement Learning (RL) with verifiable environments has emerged as a powerful approach for enhancing the reasoning capabilities of Large Language M

VIA-SD: Verification via Intra-Model Routing for Speculative Decoding

ResearchDGX agent

arXiv:2606.12243v1 Announce Type: cross Abstract: Speculative decoding (SD) addresses the high inference costs of LLMs by having lightweight drafters generate candidates for large verifiers to validat

VICX: Generalizable Robot Manipulation via Video Generation and In-Context Operator Network

Model ReleasesDGX agent

arXiv:2606.12028v1 Announce Type: new Abstract: Generalizable robot manipulation requires not only task-level reasoning over unseen scenes, but also reliable grounding of visual plans into embodiment-

VietMed-MCQ: A Consistency-Filtered Data Synthesis Framework for Vietnamese Traditional Medicine Evaluation

Model ReleasesDGX agent

arXiv:2601.03792v2 Announce Type: replace Abstract: Large Language Models (LLMs) have demonstrated remarkable proficiency in general medical domains. However, their performance significantly degrades

Vision-Aided Relative State Estimation for Approach and Landing on a Moving Platform with Inertial Measurements

ResearchDGX agent

arXiv:2512.19245v2 Announce Type: replace-cross Abstract: This paper tackles the problem of estimating the relative position, orientation, and velocity between a UAV and a planar platform undergoing a

Vision Transformers for Face Recognition Need More Registers

ResearchDGX agent

arXiv:2606.12036v1 Announce Type: new Abstract: Recent advances in Vision Transformers (ViTs) for face recognition (FR) have moved beyond the standard CLS-token paradigm. In this paradigm, a special c

Visualizing LLM Latent Space Geometry Through Dimensionality Reduction

Model ReleasesDGX agent

arXiv:2511.21594v3 Announce Type: replace Abstract: Large language models (LLMs) achieve state-of-the-art results across many natural language tasks, but their internal mechanisms remain difficult to

ViT-FREE: Efficient Face Recognition via Early Exiting and Synthetic Adaptation

SafetyDGX agent

arXiv:2606.12023v1 Announce Type: new Abstract: Vision Transformers (ViTs) have gained significant attention in computer vision and shown strong potential for face recognition (FR). However, their hig

VL-DINO: Leveraging CLIP Vision-Language Knowledge for Open-Vocabulary Object Detectio

Model ReleasesDGX agent

arXiv:2606.11546v1 Announce Type: new Abstract: Vision-language models like CLIP can provide rich semantic priors for open-vocabulary object detection. However, jointly integrating both textual and vi

VLGA: Vision-Language-Geometry-Action Models for Autonomous Driving

SafetyDGX agent

arXiv:2606.12396v1 Announce Type: new Abstract: Vision-language-action (VLA) models can describe scenes and reason about them in language, yet still struggle to ground their actions in the dense 3D wo

VOID: Defeating Unauthorized Mimicry in Latent Diffusion Models

ResearchDGX agent

arXiv:2606.12263v1 Announce Type: new Abstract: While Latent Diffusion Models (LDMs) have revolutionized visual synthesis, they are increasingly exploited for unauthorized mimicry of individuals. Exis

Weighted Random Dot Product Graphs

ResearchDGX agent

arXiv:2505.03649v4 Announce Type: replace-cross Abstract: Modeling of intricate relational patterns has become a cornerstone of contemporary statistical research and related data science fields. Netwo

What Limits Does Quantization Place on Dense Top-k Retrieval? A Theoretical Study

ResearchDGX agent

arXiv:2606.11780v1 Announce Type: cross Abstract: We establish conditions for embedding a corpus of N documents as d-dimensional vectors such that every k-subset S subseteq [N] is realizable as a resu

What Uncertainties Do We Need for Dynamical Systems?

ResearchDGX agent

arXiv:2606.11988v1 Announce Type: new Abstract: The distinction between aleatoric and epistemic uncertainty has received considerable attention in machine learning research, mainly in the context of s

When Context Returns: Toward Robust Internalization in On-Policy Distillation

SafetyDGX agent

arXiv:2606.11627v1 Announce Type: cross Abstract: Recent work has shown that on-policy distillation can internalize privileged context, such as system prompts or task hints, into a student model so th

When Do Data-Driven Systems Exhibit the Capability to Infer?

ResearchDGX agent

arXiv:2606.11769v1 Announce Type: new Abstract: The European AI Act is the first comprehensive regulation of artificial intelligence (AI), setting out extensive obligations, particularly for so-called

When Does Language Matter? Multilingual Instructions Reveal Step-wise Language Sensitivity in Vision-Language-Action Models

Model ReleasesDGX agent

arXiv:2606.11906v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have shown strong performance in language-conditioned robotic manipulation, yet their robustness to linguistic varia

When Generic Prompt Improvements Hurt: Evaluation-Driven Iteration for LLM Applications

Model ReleasesDGX agent

arXiv:2601.22025v2 Announce Type: replace-cross Abstract: Evaluating Large Language Model (LLM) applications differs from conventional software testing because outputs are probabilistic, semantically

When is Your LLM Steerable?

TutorialsDGX agent

arXiv:2606.11599v1 Announce Type: new Abstract: Activation steering offers a lightweight approach to control language models' behavior at inference time, but whether it succeeds or fails heavily depen

When More Documents Hurt RAG: Mitigating Vector Search Dilution with Domain-Scoped, Model-Agnostic Retrieval

AgentsDGX agent

arXiv:2606.11350v1 Announce Type: new Abstract: Retrieval-augmented generation degrades when scaled to large, heterogeneous document collections, where dense similarity loses discriminative power, and

When Poison Fails After Retrieval: Revisiting Corpus Poisoning under Chunking and Reranking Pipelines

ResearchDGX agent

arXiv:2606.11265v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) systems are vulnerable to corpus poisoning attacks that manipulate downstream model outputs through malicious kno

When Probing Accuracy Saturates, Fragility Resolves: A Complementary Metric for LLM Pre-Training Analysis

ResearchDGX agent

arXiv:2606.11375v1 Announce Type: cross Abstract: Standard linear probing declares a property 'encoded' when a classifier on hidden states achieves high accuracy. The protocol works well on a snapshot

When Researchers Say Mental Model/Theory of Mind of AI, What Are They Really Talking About?

SafetyDGX agent

arXiv:2510.02660v2 Announce Type: replace-cross Abstract: When researchers claim AI systems possess ToM or mental models, they are fundamentally discussing behavioral predictions and bias corrections

When Roleplaying, Do Models Believe What They Say?

Model ReleasesDGX agent

arXiv:2606.11502v1 Announce Type: cross Abstract: Language models can state that 'the Earth orbits the Sun' and, when role-playing Aristotle, assert the opposite. Recent work argues that persona adopt

Where Do Backdoors Live? A Component-Level Analysis of Backdoor Propagation in Speech Language Models

ResearchDGX agent

arXiv:2510.01157v4 Announce Type: replace Abstract: Speech language models (SLMs) are systems of systems: independent components that unite to achieve a common goal. Despite their heterogeneous nature

Which Models Are Our Models Built On? Auditing Invisible Dependencies in Modern LLMs

Model ReleasesDGX agent

arXiv:2606.12385v1 Announce Type: new Abstract: Modern LLM training pipelines increasingly rely on other models to generate data, filter corpora, judge outputs, and guide development decisions. These

Which Speech Representation Better Matches Text-Native Reasoning? A Study of Speech-Text Alignment on Frame Rate and Representation

SafetyDGX agent

arXiv:2606.12199v1 Announce Type: cross Abstract: Spoken dialogue models typically start from text LLM backbones, yet reasoning often degrades when conditioning on speech instead of text. We attribute

Why Depth Matters in Parallelizable Sequence Models: A Lie Algebraic View

ResearchDGX agent

arXiv:2603.05573v2 Announce Type: replace Abstract: Scalable sequence models, such as Transformer variants and structured state-space models, often trade expressivity power for sequence-level parallel

Wild3R: Feed-Forward 3D Gaussian Splatting from Unconstrained Sparse Photo Collection

ApplicationsDGX agent

arXiv:2606.11894v1 Announce Type: new Abstract: Feed-forward 3D Gaussian Splatting (3DGS) removes the need for time-consuming per-scene optimization required by traditional 3DGS. However, existing fee

World Model Self-Distillation: Training World Models to Solve General Tasks

Model ReleasesDGX agent

arXiv:2606.12072v1 Announce Type: new Abstract: Pretrained video generators are promising visual world models that exhibit emergent task-solving abilities; however, their reliance on detailed textual

World Pilot: Steering Vision-Language-Action Models with World-Action Priors

Model ReleasesDGX agent

arXiv:2606.12403v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models inherit semantic grounding from large-scale pretraining and perform competently across in-distribution manipulation

WorldReasoner: Evaluating Whether Language Model Agents Forecast Events with Valid Reasoning

Model ReleasesDGX agent

arXiv:2606.11816v1 Announce Type: cross Abstract: Forecasting real-world events requires language-model agents to reason under uncertainty from incomplete, time-bounded information. Yet evaluating whe

XPR: An Extensible Cross-Platform Point-Based Differentiable Renderer

ResearchDGX agent

arXiv:2606.11529v1 Announce Type: cross Abstract: Point-based differentiable rendering underpins modern 3D reconstruction, novel-view synthesis, and learning-based graphics pipelines, but developing n

10 Jun 2026

3D-CoS: A New 3D Reconstruction Paradigm Based on VLM Code Synthesis

AgentsDGX agent

arXiv:2606.10478v1 Announce Type: new Abstract: Most recent 3D reconstruction and editing systems operate on implicit and explicit representations such as NeRF, point clouds, or meshes. While these re

3SPO: State-Score-Supervised Policy Optimization for LLM Agents

SafetyDGX agent

arXiv:2606.09961v1 Announce Type: cross Abstract: Training large language models (LLMs) as autonomous agents via reinforcement learning (RL) has enabled frontier models to achieve superhuman performan

5% > 100%: Flatness Preference is All You Need for Multimodal Parameter-Efficient Fine-Tuning

Model ReleasesDGX agent

arXiv:2606.10488v1 Announce Type: new Abstract: Parameter-Efficient Fine-Tuning (PEFT) methods provide a streamlined and efficient tool for adapting large models to domain-specific multimodal downstre

A Bayesian Network Approach for Enhancing Security-Focused Decision Support Systems

TutorialsDGX agent

arXiv:2606.10782v1 Announce Type: cross Abstract: The adoption and integration of heterogeneous stacks in most of today's open-source based networks brings clear benefits like interoperability and ava

A complementary study on PlanGPT: Evaluation with defined Performance Metrics and comparison with a planner

Model ReleasesDGX agent

arXiv:2606.10489v1 Announce Type: new Abstract: Automated Planning is a subfield of Artificial Intelligence (AI) where the main objective is generating a sequence of actions, known as a plan, that hel

← Previous
1…423424425426427…1049
Next →