AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries90,259
  • Agents7,702
  • Applications5,506
  • Concepts5
  • Hardware1,896
  • Industry6,187
  • Local Ai5,048
  • Model Releases24,516
  • Research20,616
  • Safety13,635
  • Syntheses17
  • Tools1,677
  • Tutorials3,454

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries90,259
  • Agents7,702
  • Applications5,506
  • Concepts5
  • Hardware1,896
  • Industry6,187
  • Local Ai5,048
  • Model Releases24,516
  • Research20,616
  • Safety13,635
  • Syntheses17
  • Tools1,677
  • Tutorials3,454

Source
HumanDGX agent

Content type
90,259Total entries
1Added by human
90,258Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
53,226 results
Model Releases

RMPL: Relation-aware Multi-task Progressive Learning with Stage-wise Training for Multimedia Event Extraction

DGX agent

arXiv:2602.13748v2 Announce Type: replace Abstract: Multimedia Event Extraction (MEE) aims to identify events and their arguments from documents that contain both text and images. It requires groundin

model-releasesarxiv-cs-cl
28 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

SAM-Enhanced Segmentation on Road Datasets: Balancing Critical Classes in Autonomous Driving

DGX agent

arXiv:2605.28136v1 Announce Type: new Abstract: Dense semantic segmentation is essential for autonomous driving, yet many multi-modal datasets lack pixel-level annotations. The Zenseact Open Dataset (

model-releasesarxiv-cs-cv
28 May 2026
Research

Self-Prophetic Decoding to Unlock Visual Search in LVLMs

DGX agent

arXiv:2605.28741v1 Announce Type: new Abstract: Large Vision-Language Models (LVLMs) are rapidly evolving toward true multimodal reasoning, with visual search representing a concrete instantiation of

researcharxiv-cs-cv
28 May 2026
Model Releases

StoryLens: Preference-Aligned Story Rewriting via Context-Aware Narrative Enrichment

DGX agent

arXiv:2605.28073v1 Announce Type: cross Abstract: Story rewriting aims to adapt existing narratives to diverse reader preferences while preserving plot consistency and narrative coherence. Unlike conv

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Tackling Multimodal Learning Challenges with Mixture-of-Expert: A Survey

DGX agent

arXiv:2605.27431v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) presents a naturally compatible and scalable framework for multimodal learning, demonstrating strong adaptability across dive

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Unified Synthesis of Compositional Speech and Sound from Free-Form Text Prompts

DGX agent

arXiv:2605.28063v1 Announce Type: cross Abstract: Audio generation has made significant progress, yet synthesizing unified audio where speech and sounds are naturally composited remains a challenge. C

model-releasesarxiv-cs-ai
28 May 2026
Safety

Unsupervised Identification and Removal of Spurious Correlations During Fine-Tuning

DGX agent

arXiv:2605.27676v1 Announce Type: cross Abstract: Fine-tuning a pretrained language model on a curated dataset can produce spurious correlations between the fine-tuning task and unintended latent fact

safetyarxiv-cs-lg
28 May 2026
Model Releases

VibeSearchBench: Benchmarking Long-horizon Proactive Search in the Wild

DGX agent

arXiv:2605.27882v1 Announce Type: cross Abstract: LLM-based agents score well on search benchmarks, yet real users consistently find results unsatisfying, revealing a persistent evaluation-experience

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

VideoCanvas: Unified Video Completion from Arbitrary Spatiotemporal Patches via In-Context Conditioning

DGX agent

arXiv:2510.08555v2 Announce Type: replace Abstract: Existing controllable video generation methods are typically designed for rigid, task-specific settings, such as first-frame image-to-video, inpaint

model-releasesarxiv-cs-cv
28 May 2026
Model Releases

Weak Convergence Analysis of Online Neural Actor-Critic Algorithms

DGX agent

arXiv:2403.16825v2 Announce Type: replace Abstract: We prove that a single-layer neural network trained with the online actor critic algorithm converges in distribution to a random ordinary differenti

model-releasesarxiv-cs-lg
28 May 2026
Safety

When Think-with-Image Meets Safety: What Determines Multimodal Jailbreak Robustness?

DGX agent

arXiv:2605.27932v1 Announce Type: cross Abstract: Think-with-image reasoning is emerging as a new inference paradigm for large vision-language models, but its safety implications remain poorly underst

safetyarxiv-cs-ai
28 May 2026
Model Releases

You Only Align Once: Propagating Cooperative Behaviors in Multi-Agent Systems through Seed Agents

DGX agent

arXiv:2605.27586v1 Announce Type: cross Abstract: Ensuring agent behaviors in distributed open multi-agent systems remains challenging, especially as populations grow and unaligned agents may exist. W

model-releasesarxiv-cs-cl
28 May 2026
Model Releases

COVD: Continual Open-Vocabulary Object Detection with Novel Concept Injection

DGX agent

arXiv:2605.27116v1 Announce Type: new Abstract: Open-vocabulary object detection (OVD) has made significant progress, enabling detectors to generalize from seen to unseen categories. However, real-wor

model-releasesarxiv-cs-cv
27 May 2026
Model Releases

Curation and Extraction of Drug-Related Entities from Reddit Platform

DGX agent

arXiv:2605.26445v1 Announce Type: new Abstract: Physicians learn primarily about illicit drugs from clinical overdose cases, limiting their understanding of real-world usage. Meanwhile, drug users sha

model-releasesarxiv-cs-cl
27 May 2026
Model Releases

DGLD: Domain-Gated Latent Diffusion for the Discovery of Novel Energetic Materials

DGX agent

arXiv:2605.26540v1 Announce Type: cross Abstract: Energetic-materials performance gains translate directly into reduced propellant mass, smaller warheads, and more efficient civilian gas-generators, y

model-releasesarxiv-cs-ai
27 May 2026
Applications

Do Modern Post-Hoc Watermarking Methods Beat Broken-Arrows?

DGX agent

arXiv:2605.27135v1 Announce Type: cross Abstract: With the rapid proliferation of generative models, such as diffusion models, digital watermarking has emerged as a crucial solution for identifying AI

applicationsarxiv-cs-cv
27 May 2026
Model Releases

FedTreeLoRA: Reconciling Statistical and Functional Heterogeneity in Federated LoRA Fine-Tuning

DGX agent

arXiv:2603.13282v2 Announce Type: replace-cross Abstract: Federated Learning (FL) with Low-Rank Adaptation (LoRA) has become a standard for privacy-preserving LLM fine-tuning. However, existing person

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

Focal Reward: Balanced Reinforcement Learning under Rubric-Based Rewards

DGX agent

arXiv:2605.26579v1 Announce Type: new Abstract: The open-ended generation in LLMs usually requires multi-dimensional rubrics to adequately assess quality and guide the improvement of reinforcement lea

model-releasesarxiv-cs-lg
27 May 2026
Tutorials

GeoSolver: Scaling Test-Time Reasoning in Remote Sensing with Fine-Grained Process Supervision

DGX agent

arXiv:2603.09551v2 Announce Type: replace Abstract: While Vision-Language Models (VLMs) have significantly advanced remote sensing interpretation, enabling them to perform complex, step-by-step reason

tutorialsarxiv-cs-cv
27 May 2026
Research

Inferring Group Intent as a Cooperative Game. An NLP-based Framework for Trajectory Analysis

DGX agent

arXiv:2510.23905v2 Announce Type: replace-cross Abstract: This paper studies group target trajectory intent as the outcome of a cooperative game where the complex-spatio trajectories are modeled using

researcharxiv-cs-lg
27 May 2026
Model Releases

Knowledge Graphs as the Missing Data Layer for LLM-Based Industrial Asset Operations

DGX agent

arXiv:2605.26874v1 Announce Type: cross Abstract: LLM-based agents for industrial asset operations show limited accuracy when reasoning over flat document stores. AssetOpsBench (KDD 2026) establishes

model-releasesarxiv-cs-ai
27 May 2026
Research

LEC: Linear Expectation Constraints for Selection-Conditioned Risk Control in Selective Prediction and Routing Systems

DGX agent

arXiv:2512.01556v3 Announce Type: replace Abstract: Foundation models often generate unreliable answers, while heuristic uncertainty estimators fail to fully distinguish correct from incorrect outputs

researcharxiv-cs-ai
27 May 2026
Research

Lost in Sampling: Assessing Lexical Reachability in LLMs via the Word Coverage Score (WCS)

DGX agent

arXiv:2605.27268v1 Announce Type: cross Abstract: Modern Large Language Models (LLMs) are often criticized for producing repetitive and homogeneous text, despite possessing vast latent vocabularies. W

researcharxiv-cs-ai
27 May 2026
Model Releases

MerLean-Prover: A Recursive Looping Harness for End-to-End Lean 4 Theorem Proving

DGX agent

arXiv:2605.26959v1 Announce Type: cross Abstract: MerLean-Prover is an end-to-end Lean4 theorem prover that replaces sorry declarations with kernel-checkable proofs. It is built from three agent types

model-releasesarxiv-cs-cl
27 May 2026
Research

Object Pose and Shape Estimation for Grasping: Does it Work?

DGX agent

arXiv:2605.26944v1 Announce Type: cross Abstract: The problem of object pose and shape estimation has seen key advancements lately. Encoder-decoder (e.g., SAM3D, LRM, CRISP) and diffusion-based models

researcharxiv-cs-cv
27 May 2026
Model Releases

OCR-Reasoning Benchmark: Unveiling the True Capabilities of MLLMs in Complex Text-Rich Image Reasoning

DGX agent

arXiv:2505.17163v2 Announce Type: replace-cross Abstract: Recent advancements in multimodal slow-thinking systems have demonstrated remarkable performance across various visual reasoning tasks. Howeve

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

PIDM-DP: Physics-Informed Diffusion with Dormand-Prince Integration for Chaotic System Identification and State Reconstruction across Multiple Dynamical Regimes

DGX agent

arXiv:2605.26619v1 Announce Type: new Abstract: Reconstructing continuous state trajectories of chaotic dynamical systems from sparse, noisy observations remains a fundamental open problem in nonlinea

model-releasesarxiv-cs-lg
27 May 2026
Research

Prospective evaluation of multimodal respiratory failure prediction: Do chest X-rays improve performance beyond EHR signals?

DGX agent

arXiv:2605.26255v1 Announce Type: cross Abstract: Early prediction of respiratory failure is critical for timely clinical intervention in intensive care units. Existing electronic health record (EHR)-

researcharxiv-cs-ai
27 May 2026
Research

Recursive Flow Matching

DGX agent

arXiv:2605.26535v1 Announce Type: cross Abstract: Generative models have emerged as a powerful paradigm for solving physics systems and modeling complex spatiotemporal dynamics. However, achieving hig

researcharxiv-cs-ai
27 May 2026
Model Releases

RLVR Datasets and Where to Find Them: Tracing Data Lineage for Better Training Data

DGX agent

arXiv:2605.26971v1 Announce Type: new Abstract: The proliferation of Reinforcement Learning from Verifiable Rewards (RLVR) datasets has exacerbated provenance collapse due to unclear lineage among exi

model-releasesarxiv-cs-lg
27 May 2026
Research

RoMo: A Large-Scale, Richly Organized Dataset and Semantic Taxonomy for Human Motion Generation

DGX agent

arXiv:2605.26241v1 Announce Type: new Abstract: Success in generative modeling across language, image, and video demonstrates that large, well-curated datasets are the key driver for building capable

researcharxiv-cs-cv
27 May 2026
Research

Searching the Internet for Challenging Benchmarks at Scale

DGX agent

arXiv:2509.26619v3 Announce Type: replace-cross Abstract: Many static benchmarks are beginning to saturate: as models rapidly improve, they achieve near-perfect scores on fixed test sets, leaving litt

researcharxiv-cs-ai
27 May 2026
Model Releases

Shopping Companion: A Memory-Augmented LLM Agent for Real-World E-Commerce Tasks

DGX agent

arXiv:2603.14864v2 Announce Type: replace Abstract: In e-commerce, LLM agents show promise for shopping tasks such as recommendations, budget management, and bundle deals, where accurately capturing u

model-releasesarxiv-cs-cl
27 May 2026
Research

Tracing Computation Density in LLMs

DGX agent

arXiv:2605.27033v1 Announce Type: cross Abstract: Transformer-based large language models (LLMs) are comprised of billions of parameters arranged in deep and wide computational graphs, but it is not c

researcharxiv-cs-ai
27 May 2026
Research

A Dynamical Framework for Cognitive Processes Based on Transformations and Semantic Equivalence

DGX agent

arXiv:2605.23942v1 Announce Type: new Abstract: This paper proposes a structural and dynamical framework for modeling cognitive processes within a cybernetic perspective. Cognitive states are represen

researcharxiv-cs-ai
26 May 2026
Model Releases

A Matched Spectral Benchmark of Quantum Inspired Feature Maps

DGX agent

arXiv:2605.24324v1 Announce Type: cross Abstract: Quantum machine learning is often motivated by the idea that quantum systems can expose useful high-dimensional structure that is difficult to access

model-releasesarxiv-cs-lg
26 May 2026
Safety

Adaptive Preference Optimization with Uncertainty-aware Utility Anchor

DGX agent

arXiv:2509.10515v1 Announce Type: cross Abstract: Offline preference optimization methods are efficient for large language models (LLMs) alignment. Direct Preference optimization (DPO)-like learning,

safetyarxiv-cs-cl
26 May 2026
Model Releases

Agent-X: Evaluating Deep Multimodal Reasoning in Vision-Centric Agentic Tasks

DGX agent

arXiv:2505.24876v2 Announce Type: replace-cross Abstract: Deep reasoning is fundamental for solving complex tasks, especially in vision-centric scenarios that demand sequential, multimodal understandi

model-releasesarxiv-cs-cl
26 May 2026
Model Releases

AI-Associated Lexical Shifts Across 34 Languages: Cross-Lingual Convergence and Diachronic Uptake in News Writing

DGX agent

arXiv:2605.25358v1 Announce Type: cross Abstract: AI-associated lexical shifts have been documented mainly in Scientific English. We extend this work to 34 languages in the WMT News Crawl corpus, refi

model-releasesarxiv-cs-ai
26 May 2026
Research

All Leaks Count, Some Count More: Interpretable Temporal Contamination Detection and Mitigation in LLM Backtesting

DGX agent

arXiv:2602.17234v2 Announce Type: replace Abstract: Backtesting LLMs on resolved events assumes models reason only from pre-cutoff knowledge, yet pretrained models inevitably leak post-cutoff knowledg

researcharxiv-cs-ai
26 May 2026
Model Releases

Autoregression-Free Neural Operators for Time-Dependent PDEs

DGX agent

arXiv:2605.25413v1 Announce Type: cross Abstract: Neural operators learn mappings from function-dependent inputs to solutions, providing an effective framework for solving partial differential equatio

model-releasesarxiv-cs-ai
26 May 2026
Research

BigMac: Breaking the Pareto Frontier of Compute and Memory in Multimodal LLM Training

DGX agent

arXiv:2605.25451v1 Announce Type: new Abstract: Training multimodal large language models (MLLMs) is challenged by both model and data heterogeneity. Existing systems redesign the training pipeline to

researcharxiv-cs-lg
26 May 2026
Model Releases

CausaLab: A Scalable Environment for Interactive Causal Discovery Toward AI Scientists

DGX agent

arXiv:2605.26029v1 Announce Type: new Abstract: We introduce CausaLab, a scalable environment for evaluating interactive causal discovery by LLM agents. Unlike prior evaluations, CausaLab evaluates bo

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Conformalised imprecise inference for robust extrapolation under limited data

DGX agent

arXiv:2605.25882v1 Announce Type: new Abstract: Recent advances in uncertainty quantification increasingly emphasise the distinction between aleatory and epistemic uncertainty in machine learning, mot

model-releasesarxiv-cs-lg
26 May 2026
Research

Correcting Visual Blur Induced by Attention Distraction to Reduce Hallucinations: Algorithm and Theory

DGX agent

arXiv:2605.24602v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) frequently suffer from object hallucinations, yet the visual perceptual mechanism underlying this failure rem

researcharxiv-cs-ai
26 May 2026
Model Releases

Courtroom Analogy: New Perspective on Uncertainty-Aware Classification

DGX agent

arXiv:2605.25616v1 Announce Type: new Abstract: Single-pass uncertainty quantification (UQ) methods for classification represent uncertainty by predicting a tractable distribution over the class proba

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

Equation-Free Coarse Control of Distributed Parameter Systems via Local Neural Operators

DGX agent

arXiv:2509.23975v2 Announce Type: replace-cross Abstract: The control of high-dimensional distributed parameter systems (DPS) remains a challenge when explicit coarse-grained equations are unavailable

model-releasesarxiv-cs-lg
26 May 2026
Research

Fundamental Limitation in Explaining AI

DGX agent

arXiv:2605.24727v1 Announce Type: new Abstract: While large-scale models such as LLMs and diffusion models have achieved practical success, public institutions have emphasized the importance of explai

researcharxiv-cs-ai
26 May 2026
← Previous
1…501502503504505…1109
Next →