AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,419
  • Agents7,559
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,170
  • Local Ai4,934
  • Model Releases23,909
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,419
  • Agents7,559
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,170
  • Local Ai4,934
  • Model Releases23,909
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
88,419Total entries
1Added by human
88,418Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
51,935 results
Model Releases

LLMs Are Already Good Tutors: Training-Free Prompt Optimization for Pedagogical Math Tutoring

DGX agent

arXiv:2605.27088v1 Announce Type: new Abstract: Aligning LLMs for math tutoring typically requires RL-based training with multi-GPU infrastructure. We investigate whether training-free prompt optimiza

model-releasesarxiv-cs-cl
27 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

LLMs versus the Halting Problem: Characterizing Program Termination Reasoning

DGX agent

arXiv:2601.18987v5 Announce Type: replace-cross Abstract: Determining whether a program terminates is a central problem in computer science. Turing's Halting Problem established termination as undecid

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

LongAV-Compass: Towards Unified Evaluation of Minute-Scale Audio-Visual Generation Across T2AV, I2AV, and V2AV

DGX agent

arXiv:2605.26244v1 Announce Type: new Abstract: Audio-visual generation is rapidly advancing from short clips to minute-long content, while existing evaluation protocols remain largely confined to sho

model-releasesarxiv-cs-cv
27 May 2026
Model Releases

LongCat-Video-Avatar 1.5 Technical Report

DGX agent

arXiv:2605.26486v1 Announce Type: new Abstract: Despite advances in audio-driven video generation, achieving commercial-grade stability remains challenging. We present LongCat-Video-Avatar 1.5, an upg

model-releasesarxiv-cs-cv
27 May 2026
Research

MetaSICL: Adapting Audiroty LLM via Meta Speech In-Context Learning

DGX agent

arXiv:2601.18904v2 Announce Type: replace-cross Abstract: Auditory Large Language Models (LLMs) have demonstrated strong performance across a wide range of speech and audio understanding tasks. Nevert

researcharxiv-cs-ai
27 May 2026
Safety

Modernising Reinforcement Learning-Based Navigation for Embodied Semantic Scene Graph Generation

DGX agent

arXiv:2603.25415v2 Announce Type: replace Abstract: Semantic world models enable embodied agents to reason about objects, relations, and spatial context beyond purely geometric representations. In Org

safetyarxiv-cs-ai
27 May 2026
Model Releases

Not All Disagreement Is Learnable: Token Teachability in On-Policy Distillation

DGX agent

arXiv:2605.26844v1 Announce Type: new Abstract: On-policy distillation (OPD) trains a student on its own rollouts with token-level teacher supervision. Recent selective OPD methods exploit the non-uni

model-releasesarxiv-cs-lg
27 May 2026
Model Releases

PilotTTS: A Disciplined Modular Recipe for Competitive Speech Synthesis

DGX agent

arXiv:2605.27258v1 Announce Type: cross Abstract: Building state-of-the-art text-to-speech (TTS) systems typically demands millions of hours of proprietary data and complex multi-stage architectures,

model-releasesarxiv-cs-ai
27 May 2026
Research

PinPoint: Prompting with Informative Interior Points

DGX agent

arXiv:2605.26689v1 Announce Type: cross Abstract: Modern referring image segmentation pipelines couple a vision-language model (VLM) for grounding with a promptable segmenter such as the Segment Anyth

researcharxiv-cs-cl
27 May 2026
Local Ai

Plan Then Action:High-Level Planning Guidance Reinforcement Learning for LLM Reasoning

DGX agent

arXiv:2510.01833v2 Announce Type: replace Abstract: Large language models (LLMs) demonstrate strong reasoning abilities via Chain-of-Thought (CoT), but their token-level generation encourages local de

local-aiarxiv-cs-ai
27 May 2026
Model Releases

Provably Communication-Efficient and Privacy-Preserving Federated Graph Neural Networks

DGX agent

arXiv:2605.26243v1 Announce Type: new Abstract: Graph neural networks (GNNs) achieve strong performance on relational data, but real-world graphs are often distributed across organizations that cannot

model-releasesarxiv-cs-lg
27 May 2026
Applications

Reasoning Depth and Environment Complexity: A Controlled Study of RLVR Data Allocation across Logical Reasoning Tasks

DGX agent

arXiv:2605.26934v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has become central to post-training reasoning models, yet a key limitation of existing studies i

applicationsarxiv-cs-ai
27 May 2026
Model Releases

Receipt Replay OOD: A Small Benchmark for Screen Replay Detection Under Domain Shift

DGX agent

arXiv:2605.26855v1 Announce Type: new Abstract: Public datasets such as DLC-2021, SynID, and KID34K have significantly contributed to research on presentation attack detection for identity documents,

model-releasesarxiv-cs-cv
27 May 2026
Model Releases

Resolving Ambiguity in Composed Image Retrieval via Calibrated Interaction

DGX agent

arXiv:2605.24634v2 Announce Type: replace Abstract: Composed image retrieval (CIR) searches a corpus with a reference image and a text describing how to modify it. Despite rapid progress from triplet-

model-releasesarxiv-cs-cv
27 May 2026
Model Releases

ScientistOne: Towards Human-Level Autonomous Research via Chain-of-Evidence

DGX agent

arXiv:2605.26340v1 Announce Type: new Abstract: Autonomous research agents produce competitive solutions and professional-looking manuscripts, yet their outputs contain verifiability failures undetect

model-releasesarxiv-cs-ai
27 May 2026
Safety

Securing Multi-Agent Systems Against Corruptions via Node Contribution Backpropagation

DGX agent

arXiv:2510.19420v2 Announce Type: replace-cross Abstract: Multi-Agent Systems (MAS) have become a prevalent paradigm for Large Language Model (LLM) applications. However, the complex multi-agent desig

safetyarxiv-cs-ai
27 May 2026
Agents

Self-signals Driven Multi-LLM Debate for Efficient and Accurate Reasoning

DGX agent

arXiv:2510.06843v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have exhibited impressive capabilities across diverse application domains. Recent work has explored Multi-LLM Age

agentsarxiv-cs-ai
27 May 2026
Model Releases

Sentinel: Embodied Cooperative Spatial Reasoning and Planning

DGX agent

arXiv:2605.26239v1 Announce Type: new Abstract: In this work, we study Cooperative Spatial Intelligence, the ability of decentralized embodied agents to coordinate effectively under dynamic environmen

model-releasesarxiv-cs-cv
27 May 2026
Local Ai

Separate Aggregation of Split Network for Personalized Federated Learning

DGX agent

arXiv:2605.26571v1 Announce Type: new Abstract: Federated learning enables collaborative model training without sharing raw data, but its performance can degrade substantially under heterogeneous clie

local-aiarxiv-cs-lg
27 May 2026
Model Releases

Shedding Light on Dark Matter at the LHC with Machine Learning

DGX agent

arXiv:2509.15121v2 Announce Type: replace-cross Abstract: We investigate a WIMP dark matter (DM) candidate in the form of a singlino-dominated lightest supersymmetric particle (LSP) within the Z_3-sym

model-releasesarxiv-cs-lg
27 May 2026
Model Releases

SteelDS: A High-Resolution Video Dataset of E40 Steel Scrap for Object Detection and Instance Segmentation

DGX agent

arXiv:2605.26682v1 Announce Type: cross Abstract: This dataset provides high-resolution, annotated video sequences of shredded E40-grade steel and copper scrap on a conveyor belt. Captured in a contro

model-releasesarxiv-cs-cv
27 May 2026
Safety

TAGRPO: Boosting GRPO on Image-to-Video Generation with Direct Trajectory Alignment

DGX agent

arXiv:2601.05729v2 Announce Type: replace Abstract: Recent studies have demonstrated the efficacy of integrating Group Relative Policy Optimization (GRPO) into flow matching models, particularly for t

safetyarxiv-cs-cv
27 May 2026
Model Releases

Temporal Simultaneity Predicts Annotation Quality in Sentiment Corpora

DGX agent

arXiv:2605.27239v1 Announce Type: new Abstract: Annotation quality is difficult to sustain when campaigns span weeks or months with small annotator pools. We present a Setswana sentiment dataset of 3,

model-releasesarxiv-cs-cl
27 May 2026
Model Releases

Towards Error-Free EHRs: Reasoning-Intensive Consistency Verification Between Clinical Notes and Structured Tables in Electronic Health Records

DGX agent

arXiv:2605.26463v1 Announce Type: cross Abstract: Data consistency between unstructured clinical notes and structured tables in Electronic Health Records (EHRs) is essential for patient safety and cli

model-releasesarxiv-cs-ai
27 May 2026
Local Ai

Towards Interpretable Federated Learning

DGX agent

arXiv:2302.13473v2 Announce Type: replace Abstract: Federated learning (FL) enables multiple data owners to build machine learning models collaboratively without exposing their private local data. In

local-aiarxiv-cs-lg
27 May 2026
Model Releases

Traceable Knowledge Graph Reasoning Enables LLM-Assisted Decision Support for Industrial VOCs in the Steel Industry

DGX agent

arXiv:2605.27071v1 Announce Type: new Abstract: Key knowledge for steel-industry volatile organic compounds (VOCs) governance is scattered across unstructured scientific literature, making it difficul

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

Trust Region Q Adjoint Matching

DGX agent

arXiv:2605.27079v1 Announce Type: cross Abstract: Off-policy reinforcement learning of pretrained flow policies remains challenging due to the instability of optimization arising from the multi-step s

model-releasesarxiv-cs-ai
27 May 2026
Safety

UCPO: Uncertainty-Aware Policy Optimization

DGX agent

arXiv:2601.22648v2 Announce Type: replace Abstract: The key to building trustworthy large language models (LLMs) lies in endowing them with inherent uncertainty expression capabilities, thereby mitiga

safetyarxiv-cs-ai
27 May 2026
Research

Uncertainty-Aware Budget Allocation for Adaptive Test-Time Reasoning

DGX agent

arXiv:2605.26849v1 Announce Type: new Abstract: Sampling multiple responses improves language model reasoning, but uniform compute allocation is inefficient: easy questions are over-sampled while hard

researcharxiv-cs-cl
27 May 2026
Model Releases

Underwater360: Reconstructing Underwater Scenes from Panoramic Images with Omnidirectional Gaussian Splatting

DGX agent

arXiv:2605.26447v1 Announce Type: new Abstract: Underwater scene reconstruction is essential for immersive exploration of aquatic environments, yet remains challenging due to complex participating-med

model-releasesarxiv-cs-cv
27 May 2026
Model Releases

Variational Inference for Evidential Deep Learning

DGX agent

arXiv:2605.26477v1 Announce Type: new Abstract: While Deep Neural Networks (DNNs) achieve remarkable performance, their tendency to produce overconfident predictions. Evidential Deep Learning (EDL) mi

model-releasesarxiv-cs-lg
27 May 2026
Model Releases

VISTA: An End-to-End Benchmark for Visual Spec-to-Web-App Coding Agents

DGX agent

arXiv:2605.26144v1 Announce Type: cross Abstract: We present VISTA (VIsual Spec-To-App Benchmark), a benchmark for evaluating the end-to-end web-app generation capabilities of LLM-based agents. Unlike

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

What Molecular Structure Cannot Tell Us: A Taxonomy of Explainability Gaps in GNN-Based Drug Toxicity Prediction

DGX agent

arXiv:2605.26183v1 Announce Type: cross Abstract: Graph Neural Networks (GNNs) have emerged as a structurally natural approach for molecular toxicity prediction, operating directly on atomic connectiv

model-releasesarxiv-cs-lg
27 May 2026
Model Releases

Zero-Shot Object Re-Identification in Egocentric Kitchen Videos via Multi-Stage SAM3 Feature Fusion

DGX agent

arXiv:2605.26383v1 Announce Type: new Abstract: Object re-identification (ReID) in egocentric kitchen videos is challenging due to rapid viewpoint changes, frequent occlusions, cluttered scenes, and l

model-releasesarxiv-cs-cv
27 May 2026
Model Releases

A Controlled Synthetic Benchmark for Educational Aspect-Based Sentiment Analysis

DGX agent

arXiv:2605.25502v1 Announce Type: cross Abstract: Educational aspect-based sentiment analysis (ABSA) can support course improvement, but public aspect-labeled student feedback remains scarce because e

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

A lift for input-convex neural network training

DGX agent

arXiv:2605.24274v1 Announce Type: new Abstract: Input-convex neural networks (ICNNs) are widely used for log-concave density estimation, convex-potential normalizing flows, optimal transport, and tran

model-releasesarxiv-cs-lg
26 May 2026
Applications

A Lightweight Hybrid Transformer-CRF Architecture for Multi-Type Bangla Medical Entity Recognition

DGX agent

arXiv:2605.25463v1 Announce Type: new Abstract: MedER refers to the identification of medical entities. It is crucial for extracting structured clinical information from unstructured medical text. Man

applicationsarxiv-cs-cl
26 May 2026
Model Releases

A Two-Phase Stability Study of LLM Judges and Bar Council Examiners on Thai Bar-Exam Free-Form Essays

DGX agent

arXiv:2605.25652v1 Announce Type: new Abstract: Free-form legal essay evaluation in NLP treats expert inter-rater stability as a single ceiling number, and treats LLM-judge agreement with that ceiling

model-releasesarxiv-cs-cl
26 May 2026
Model Releases

Action-Prior Denoising for Smooth Real-Time Chunking

DGX agent

arXiv:2605.25537v1 Announce Type: new Abstract: Real-time chunking (RTC) lets chunked action policies operate under inference delay by conditioning a newly generated action chunk on actions already co

model-releasesarxiv-cs-ro
26 May 2026
Model Releases

AI Cartography: Mapping the Latent Landscape of AI Benchmark Ecosystems

DGX agent

arXiv:2605.25272v1 Announce Type: new Abstract: While aggregate leaderboard scores drive AI development, they contain substantial measurement noise whose sources and magnitudes remain unquantified, ma

model-releasesarxiv-cs-ai
26 May 2026
Research

Algometrics: Forecasting Under Algorithmic Feedback

DGX agent

arXiv:2605.23978v1 Announce Type: new Abstract: In algorithmic markets, predictive models become part of the data-generating process they aim to forecast. Once their outputs are converted into trades,

researcharxiv-cs-lg
26 May 2026
Tutorials

An Empirical Evaluation of LLM-Generated Code Security Across Prompting Methods

DGX agent

arXiv:2605.24298v1 Announce Type: cross Abstract: The growing use of Large Language Models (LLMs) for automated code generation has enhanced software development efficiency, but often at the cost of s

tutorialsarxiv-cs-ai
26 May 2026
Applications

ASTRO: Adaptive Spatio-Temporal Reinforcement Optimization for GNN Powered Anomly Detection in Cyber Physical Systems

DGX agent

arXiv:2605.25135v1 Announce Type: cross Abstract: Anomaly detection in Industrial Internet of Things (IIoT) environments is essential to protect the Industrial Control Systems (ICS) and Cyber-Physical

applicationsarxiv-cs-ai
26 May 2026
Model Releases

AuthTrace: Diagnosing Evidence Construction in Thematically Dense Single-Author Corpora

DGX agent

arXiv:2605.25382v1 Announce Type: new Abstract: Evidence construction systems--chunk retrieval, agent memory, knowledge-graph traversal, and thematic indexing--are evaluated on separate benchmarks wit

model-releasesarxiv-cs-cl
26 May 2026
Model Releases

Benchmarking and Learning Real-World Customer Service Dialogue

DGX agent

arXiv:2510.22143v3 Announce Type: replace Abstract: Existing benchmarks and training pipelines for industrial intelligent customer service (ICS) remain misaligned with real-world dialogue requirements

model-releasesarxiv-cs-cl
26 May 2026
Research

Binding Visual Features Point by Point

DGX agent

arXiv:2605.25427v1 Announce Type: cross Abstract: Despite success on standard benchmarks, vision language models display persistent failures on tasks involving processing of multi-object scenes, inclu

researcharxiv-cs-ai
26 May 2026
Research

Bridging the Semantic-Action Gap in Visual Token Pruning for Efficient VLA Inference

DGX agent

arXiv:2511.16449v4 Announce Type: replace-cross Abstract: Vision-Language-Action (VLA) models have shown great potential for embodied AI by integrating visual perception, language understanding, and a

researcharxiv-cs-ai
26 May 2026
Agents

CausalFlow: Causal Attribution and Counterfactual Repair for LLM Agent Failures

DGX agent

arXiv:2605.25338v1 Announce Type: cross Abstract: Large language model (LLM) agents frequently fail on multi-step tasks involving reasoning, tool use, and environment interaction. While such failures

agentsarxiv-cs-ai
26 May 2026
← Previous
1…574575576577578…1082
Next →