AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,457
  • Agents7,560
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,171
  • Local Ai4,939
  • Model Releases23,938
  • Research20,125
  • Safety13,372
  • Syntheses17
  • Tools1,677
  • Tutorials3,404

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,457
  • Agents7,560
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,171
  • Local Ai4,939
  • Model Releases23,938
  • Research20,125
  • Safety13,372
  • Syntheses17
  • Tools1,677
  • Tutorials3,404

Source
Human
88,457Total entries
1Added by human
88,456Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
62,897 results
Model Releases

What Demonstration Curation Metrics Do to Your Policy

DGX agent

arXiv:2606.10229v1 Announce Type: cross Abstract: We study whether demonstration-curation metrics that detect defective training episodes also improve the downstream behavior-cloning policy that train

model-releasesarxiv-cs-lg
10 Jun 2026
Research

What Do Deepfake Speech Detectors Actually Hear?

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2606.10912v1 Announce Type: cross Abstract: Deepfake speech detectors often output a single score without explaining why an audio sample is flagged, where in the signal the evidence lies, or wha

researcharxiv-cs-ai
10 Jun 2026
Model Releases

What Fits (Into Few Tokens) Doesn't Overfit: Compression and Generalization in ML Research Agents

DGX agent

arXiv:2606.11045v1 Announce Type: new Abstract: Reusing a held-out benchmark adaptively should, in principle, invite overfitting. Yet benchmark-driven machine learning (ML) has produced surprisingly l

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

What makes a harness a harness: necessary and sufficient conditions for an agent harness

DGX agent

arXiv:2606.10106v1 Announce Type: cross Abstract: The term agent harness now circulates widely in software engineering with generative artificial intelligence. It names the layer that wraps a language

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

What Matters in Orchestrating Robot Policies: A Systematic Study of Hierarchical VLA Agents

DGX agent

arXiv:2606.10267v1 Announce Type: cross Abstract: Hierarchical vision-language-action (Hi-VLA) systems have emerged as a promising paradigm for complex robot manipulation, by using high-level VLM plan

model-releasesarxiv-cs-ai
10 Jun 2026
Research

What Really Matters for Table LLMs? A Meta-Evaluation of Model and Data Effects

DGX agent

arXiv:2501.14717v2 Announce Type: replace Abstract: Table modeling has progressed for decades. In this work, we revisit this trajectory and highlight emerging challenges in the LLM era, particularly t

researcharxiv-cs-cl
10 Jun 2026
Safety

What Should a Skill Remember? Quality--Cost Trade-offs in Cost-Aware Skill Rewriting for Language Model Agents

DGX agent

arXiv:2606.09421v2 Announce Type: replace Abstract: Large language model agents increasingly rely on skills: reusable procedural documents encoding workflows, tool use, implementation patterns, valida

safetyarxiv-cs-cl
10 Jun 2026
Agents

What Spatial Memory Must Store: Occlusion as the Test for Language-Agent Memory

DGX agent

arXiv:2606.10299v1 Announce Type: new Abstract: Language-agent 'memory palace' systems anchor each memory to a world coordinate, on the intuition that geometry adds something text cannot. We make that

agentsarxiv-cs-ai
10 Jun 2026
Local Ai

When Attribution Patching Lies: Diagnosis and a Second-Order Correction

DGX agent

arXiv:2606.09899v1 Announce Type: cross Abstract: A central goal of mechanistic interpretability is to identify which internal components causally drive a language model's behavior. Because these impo

local-aiarxiv-cs-ai
10 Jun 2026
Model Releases

When Design Rules Break: Benchmark Composition Determines Whether Label Informativeness Predicts GNN Aggregator Choice

DGX agent

arXiv:2606.10249v1 Announce Type: new Abstract: We examine whether graph neural network (GNN) design rules generalize across benchmark families by studying aggregator selection (sum, mean, max) on 24

model-releasesarxiv-cs-lg
10 Jun 2026
Safety

When Distance Distracts: Representation Distance Bias in BT-Loss for Reward Models

DGX agent

arXiv:2512.06343v3 Announce Type: replace-cross Abstract: Reward models are central to Large Language Model (LLM) alignment within the framework of RLHF. The standard objective used in reward modeling

safetyarxiv-cs-ai
10 Jun 2026
Model Releases

When Do Autoregressive Sequence Models Forecast Physical Wavefields? A Controlled Study on Synthetic Seismograms

DGX agent

arXiv:2606.10868v1 Announce Type: new Abstract: Long-horizon autoregressive forecasting of oscillatory physical signals, such as seismograms, gravitational-wave strain, and similar wavefields is limit

model-releasesarxiv-cs-lg
10 Jun 2026
Research

When Metrics Disagree: A Meta-Analysis of Knowledge-Graph-Completion Model Benchmarking

DGX agent

arXiv:2606.10287v1 Announce Type: cross Abstract: Evaluating Knowledge Graph Completion (KGC) models remains challenging because standard assessment relies on isolated rank-based metrics such as MRR,

researcharxiv-cs-cl
10 Jun 2026
Model Releases

When RL Fails after SFT: Rejuvenating Model Plasticity for Robust SFT-to-RL Handoff

DGX agent

arXiv:2606.09932v1 Announce Type: cross Abstract: Supervised Fine-Tuning (SFT) followed by Reinforcement Learning (RL) has become a standard pipeline for Large Language Model (LLM) post-training. SFT

model-releasesarxiv-cs-ai
10 Jun 2026
Safety

When the Chain of Thought Knows Better: Failure Modes in Multi-Turn Reasoning Models

DGX agent

arXiv:2606.10740v1 Announce Type: new Abstract: Failures in multi-turn reasoning models are largely invisible to terminal-score evaluation. A model can lock onto an unsafe stance early in a long dialo

safetyarxiv-cs-ai
10 Jun 2026
Safety

When to Align, When to Predict: A Phase Diagram for Multimodal Learning

DGX agent

arXiv:2606.11190v1 Announce Type: new Abstract: Cross-modal alignment (CA) and cross-modal prediction (CP) are the dominant paradigms for multimodal representation learning, yet there is no systematic

safetyarxiv-cs-lg
10 Jun 2026
Research

Where You Inject Diversity Matters: A Unified Framework for Diverse Generation

DGX agent

arXiv:2606.10302v1 Announce Type: new Abstract: Open-ended generation tasks often require a set of meaningfully different outputs, yet large language models often produce similar generations. Existing

researcharxiv-cs-cl
10 Jun 2026
Research

Which LoRA? An Empirical Study on the Effectiveness of LoRA Techniques During Multilingual Instruction Tuning

DGX agent

arXiv:2606.10428v1 Announce Type: new Abstract: We investigate whether commonly available LoRA variants have an advantage over basic LoRA in multilingual instruction tuning. Experiments involving LoRA

researcharxiv-cs-cl
10 Jun 2026
Research

Whisfusion: Parallel ASR Decoding with Masked Diffusion

DGX agent

arXiv:2508.07048v2 Announce Type: replace-cross Abstract: Autoregressive (AR) encoder-decoder models dominate high-quality multilingual ASR, but their left-to-right decoders make inference latency sca

researcharxiv-cs-ai
10 Jun 2026
Research

Whisper-GPT -- Continuous Discrete Hybrid Representation Language Models For Speech And Music

DGX agent

arXiv:2412.11449v2 Announce Type: replace-cross Abstract: We propose WHISPER-GPT: A generative large language model (LLM) for speech and music that allows us to work with continuous audio representati

researcharxiv-cs-ai
10 Jun 2026
Model Releases

Who Brought Easter Eggs to Eid? Auditing Cultural Translation of Math Word Problems Across Diverse Languages and Regions

DGX agent

arXiv:2606.11009v1 Announce Type: new Abstract: Large language models are increasingly used to adapt math word problems for personalized learning at scale, but it remains an open question whether thos

model-releasesarxiv-cs-cl
10 Jun 2026
Research

Who Wrote the Book? Detecting and Attributing LLM Ghostwriters

DGX agent

arXiv:2603.28054v2 Announce Type: replace Abstract: In this paper, we introduce GhostWriteBench, a dataset for LLM authorship attribution. It comprises long-form texts (50K+ words per book) generated

researcharxiv-cs-cl
10 Jun 2026
Model Releases

WHU-Infra3D: A Full-stack Multi-modal Dataset and Benchmark for 3D Roadside Infrastructure Inventory

DGX agent

arXiv:2606.09882v1 Announce Type: new Abstract: The paradigm of digital twin cities is shifting from coarse visual mapping toward more precise and actionable digitization of urban assets. However, exi

model-releasesarxiv-cs-cv
10 Jun 2026
Model Releases

Workflow-GYM: Towards Long-Horizon Evaluation of Computer-use Agentic tasks in Real-World Professional Fields

DGX agent

arXiv:2606.11042v1 Announce Type: new Abstract: Recent years have witnessed the rapid evolution of AI agents toward handling increasingly complex, real-world tasks. However, existing benchmarks rarely

model-releasesarxiv-cs-ai
10 Jun 2026
Research

WorldKernel: A World Model is the Coupling Kernel of Admissible Possible Worlds

DGX agent

arXiv:2606.10934v1 Announce Type: new Abstract: A common assumption holds that enough observational and interventional data, given to a strong enough predictor, suffices. We report a failure mode that

researcharxiv-cs-ai
10 Jun 2026
Model Releases

WorldOlympiad: Can Your World Model Survive a Triathlon?

DGX agent

arXiv:2606.11129v1 Announce Type: new Abstract: We introduce WorldOlympiad, a benchmark for diagnosing video-based world models across physical faithfulness, geometric consistency, and interaction fid

model-releasesarxiv-cs-cv
10 Jun 2026
Research

WorldPlay: Towards Long-Term Geometric Consistency for Real-Time Interactive World Modeling

DGX agent

arXiv:2512.14614v2 Announce Type: replace Abstract: This paper presents WorldPlay, a streaming video diffusion model that enables real-time, interactive world modeling with long-term geometric consist

researcharxiv-cs-cv
10 Jun 2026
Model Releases

XtrAIn: Training-Guided Occlusion for Feature Attribution

DGX agent

arXiv:2606.10877v1 Announce Type: cross Abstract: Occlusion-based attribution methods provide an intuitive way to estimate feature importance by perturbing input features and measuring the resulting c

model-releasesarxiv-cs-cv
10 Jun 2026
Safety

YUBI: Yielding Universal Bidigital Interface for Bimanual Dexterous Manipulation at Scale

DGX agent

arXiv:2606.10244v1 Announce Type: cross Abstract: We introduce Yielding Universal Bidigital Interface (YUBI), a finger-aligned gripper designed to enable intuitive, ergonomic, and scalable data collec

safetyarxiv-cs-ai
10 Jun 2026
Research

ZODS-RS -- Zero-training Oriented Detection & Segmentation for Remote Sensing

DGX agent

arXiv:2606.10769v1 Announce Type: new Abstract: Remote-sensing and UAV applications need models that generalize across platforms and viewpoints without task-specific training. Yet training-free pipeli

researcharxiv-cs-cv
10 Jun 2026
Research

3D Oral Modelling with Improved Vertex Distribution Using Matching-Based Learning

DGX agent

arXiv:2606.07907v1 Announce Type: cross Abstract: In our previous work, a deep learning-based framework for 3D intraoral reconstruction was proposed. The model directly predicts explicit 3D point clou

researcharxiv-cs-ai
9 Jun 2026
Safety

6G Empowering Future Robotics: A Vision for Next-Generation Autonomous Systems

DGX agent

arXiv:2602.12246v2 Announce Type: replace-cross Abstract: The convergence of robotics and next-generation communication is a critical driver of technological advancement. As the world transitions from

safetyarxiv-cs-ro
9 Jun 2026
Model Releases

A Baseline Study and Benchmark for Few-Shot Open-Set Action Recognition with Feature Residual Discrimination

DGX agent

arXiv:2603.04125v2 Announce Type: replace Abstract: Few-Shot Action Recognition (FS-AR) has shown promising results but is often limited by a closed-set assumption that fails in real-world open-set sc

model-releasesarxiv-cs-cv
9 Jun 2026
Research

A Camera-Native Talking-Head Video Dataset for Various Computer Vision Tasks

DGX agent

arXiv:2603.26763v2 Announce Type: replace Abstract: Talking-head videos constitute a predominant content type in real-time communication, yet publicly available datasets for video processing research

researcharxiv-cs-cv
9 Jun 2026
Agents

A case study of evaluating AI agents on a neuroscience data-to-discovery pipeline

DGX agent

arXiv:2606.07718v1 Announce Type: new Abstract: Agentic AI tools offer a promising path to automating software development bottlenecks in scientific research pipelines, particularly for stages that ta

agentsarxiv-cs-ai
9 Jun 2026
Model Releases

A Comparative Study of Student Perspectives on Technical Writing Feedback Quality: Evaluating LLMs, SLMs, and Humans in Computer Science Topics

DGX agent

arXiv:2601.11541v2 Announce Type: replace-cross Abstract: To address the scalability of feedback in computer science while mitigating the privacy and cost limitations of commercial Large Language Mode

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

A Comparison of SSL-Based Feature Extractors and Back-End Classifiers for Spoofing Detection: A Multi-Corpus Training and Cross-Linguistic Analysis

DGX agent

arXiv:2606.08669v1 Announce Type: cross Abstract: Voice biometric systems face growing threats from spoofing attacks, yet the evaluation of detection models remains inconsistent across datasets. To in

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

A Dataset for Dynamic Human Preferences for Vision Language Models

DGX agent

arXiv:2606.07653v1 Announce Type: cross Abstract: Given the increased adoption of Vision Language Models (VLMs) in human-interactive settings, it is important that we evaluate how well these models ca

model-releasesarxiv-cs-ai
9 Jun 2026
Safety

A Finetuned SpeechLLM for Joint Multi-Granular L2 Assessment and Natural-Language Rationales

DGX agent

arXiv:2606.09470v1 Announce Type: cross Abstract: Automated L2 speech assessment can assign proficiency labels, but often lacks interpretability. We propose a rubric-guided SpeechLLM for multi-aspect,

safetyarxiv-cs-ai
9 Jun 2026
Model Releases

A Framework for Evaluating and Benchmarking Concept Drift Detection Methods

DGX agent

arXiv:2606.07789v1 Announce Type: new Abstract: Data stream mining is fundamentally challenged by concept drift, where distributional changes can degrade model performance. Despite the proliferation o

model-releasesarxiv-cs-lg
9 Jun 2026
Research

A generalizable 3D framework and model for self-supervised learning in medical imaging

DGX agent

arXiv:2501.11755v2 Announce Type: replace-cross Abstract: Current self-supervised learning methods for 3D medical imaging rely on simple pretext formulations and organ- or modality-specific datasets,

researcharxiv-cs-cv
9 Jun 2026
Applications

A Geometric Framework for Absolute Pose and Velocity Estimation with Event Cameras

DGX agent

arXiv:2606.09139v1 Announce Type: new Abstract: Despite the rapid advancements in event-based motion estimation, current geometric methods primarily focus on velocity estimation. However, absolute pos

applicationsarxiv-cs-cv
9 Jun 2026
Research

A Geometric Measure of Linear Separability for Neural Representations

DGX agent

arXiv:2606.08721v1 Announce Type: new Abstract: Modern neural classifiers commonly rely on linear readouts, yet predictive metrics alone do not characterize the class-wise geometry of the representati

researcharxiv-cs-lg
9 Jun 2026
Research

A Geometric Theory of Cognition for Machine Intelligence

DGX agent

arXiv:2512.12225v3 Announce Type: replace Abstract: Developing artificial agents that unify representation, memory, adaptation, and prediction remains a fundamental challenge in artificial intelligenc

researcharxiv-cs-ai
9 Jun 2026
Safety

A Geometric Unification of Concept Learning with Concept Cones

DGX agent

arXiv:2512.07355v2 Announce Type: replace Abstract: Two traditions of interpretability have evolved side by side but seldom spoken to each other: Concept Bottleneck Models (CBMs), which prescribe what

safetyarxiv-cs-ai
9 Jun 2026
Local Ai

A Geometry-Aware Triplane Field Network for Vehicle Aerodynamic Prediction

DGX agent

arXiv:2606.07724v1 Announce Type: new Abstract: High-fidelity computational fluid dynamics (CFD) is crucial to vehicle aerodynamic analysis, but its cost still constrains early-stage design exploratio

local-aiarxiv-cs-lg
9 Jun 2026
Research

A Graphop Analysis of Graph Neural Networks on Sparse Graphs: Generalization and Universal Approximation

DGX agent

arXiv:2602.08785v2 Announce Type: replace Abstract: Generalization and approximation capabilities of message passing graph neural networks (MPNNs) are often studied by defining a compact metric on a s

researcharxiv-cs-lg
9 Jun 2026
Research

A Hierarchical Feature Engineering Framework for Automated Classification of Phonotraumatic and Non-Phonotraumatic Vocal Hyperfunction

DGX agent

arXiv:2606.07673v1 Announce Type: cross Abstract: Ambulatory neck-surface acceleration enables non-invasive monitoring of vocal hyperfunction, yet robust biomarkers for its subtypes remain limited. Th

researcharxiv-cs-ai
9 Jun 2026
← Previous
1…545546547548549…1311
Next →