AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,965
  • Agents7,446
  • Applications5,325
  • Concepts5
  • Hardware1,798
  • Industry6,131
  • Local Ai4,857
  • Model Releases23,360
  • Research19,834
  • Safety13,174
  • Syntheses17
  • Tools1,670
  • Tutorials3,348

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,965
  • Agents7,446
  • Applications5,325
  • Concepts5
  • Hardware1,798
  • Industry6,131
  • Local Ai4,857
  • Model Releases23,360
  • Research19,834
  • Safety13,174
  • Syntheses17
  • Tools1,670
  • Tutorials3,348

Source
HumanDGX agent

Content type
86,965Total entries
1Added by human
86,964Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
51,106 results
Model Releases

Mean-Field PhiBE: Continuous-Time Mean-Field Reinforcement Learning from Discrete-Time Data

DGX agent

arXiv:2606.26498v1 Announce Type: cross Abstract: This paper addresses model-free continuous-time mean-field control in a setting where the population dynamics evolve continuously according to an unkn

model-releasesarxiv-cs-lg
26 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

OI-Bench: An Option Injection Benchmark for Evaluating LLM Susceptibility to Directive Interference

DGX agent

arXiv:2601.13300v2 Announce Type: replace Abstract: Benchmarking large language models (LLMs) is critical for understanding their capabilities, limitations, and robustness. In addition to interface ar

model-releasesarxiv-cs-cl
26 Jun 2026
Model Releases

Perception, Verdict, and Evolution: Hindsight-Driven Self-Refining Forensics Agent for AI-Generated Image Detection

DGX agent

arXiv:2606.26552v1 Announce Type: cross Abstract: The rapid advancement of generative models presents a significant challenge to existing deepfake detection methods, particularly given the widespread

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

PhyEditBench: A Real-World Multi-Stage Benchmark for Physics-Aware Image Editing

DGX agent

arXiv:2606.26551v1 Announce Type: new Abstract: While instruction-based image editing, enabled by multi-modal generative models, has advanced significantly, existing benchmarks lack a comprehensive ev

model-releasesarxiv-cs-cv
26 Jun 2026
Model Releases

Spurious Rewards Paradox: Mechanistically Understanding How RLVR Activates Memorization Shortcuts in LLMs

DGX agent

arXiv:2601.11061v2 Announce Type: replace-cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) is highly effective for enhancing LLM reasoning, yet recent evidence shows models like Q

model-releasesarxiv-cs-cl
26 Jun 2026
Model Releases

What We are Missing in Multimodal LLM Evaluation?

DGX agent

arXiv:2606.26348v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) can process diverse inputs, e.g., text, images, audio, and video, and generate textual responses. While their c

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

2K Retrofit: Entropy-Guided Efficient Sparse Refinement for High-Resolution 3D Geometry Prediction

DGX agent

arXiv:2603.19964v3 Announce Type: replace Abstract: High-resolution geometric prediction is essential for robust perception in autonomous driving, robotics, and AR/MR, but current foundation models ar

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

BOFA: Bridge-Layer Orthogonal Low-Rank Fusion for CLIP-Based Class-Incremental Learning

DGX agent

arXiv:2511.11421v2 Announce Type: replace Abstract: Class-Incremental Learning (CIL) aims to continually learn new categories without forgetting previously acquired knowledge. Vision-language models s

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

C3-Bench: A Context-Aware Change Captioning Benchmark

DGX agent

arXiv:2606.25445v1 Announce Type: new Abstract: While Change Captioning systems have garnered substantial attention to respond to our evolving world, their true performance on diverse real-world chang

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

CausalRAG2: Hierarchical Causal Knowledge Graph Design for RAG

DGX agent

arXiv:2602.05143v2 Announce Type: replace Abstract: Retrieval augmented generation (RAG) has enhanced large language models by enabling access to external knowledge, with graph-based RAG emerging as a

model-releasesarxiv-cs-ai
25 Jun 2026
Model Releases

Curvature-Guided Mixing for MLLM Adaptation

DGX agent

arXiv:2606.24963v1 Announce Type: new Abstract: Fine-tuning Multimodal Large Language Models (MLLMs) on specialized tasks often leads to catastrophic forgetting of their general capabilities. Existing

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

Dream at SemEval-2026 Task 13: SALSA for Single-Pass Machine-Generated Code Detection

DGX agent

arXiv:2606.25102v1 Announce Type: new Abstract: Large language models have transformed code generation, raising concerns around authorship, assessment integrity, and software trust. SemEval-2026 Task

model-releasesarxiv-cs-cl
25 Jun 2026
Model Releases

Evidence for feature-specific error correction in LLMs

DGX agent

arXiv:2606.24964v1 Announce Type: new Abstract: Understanding the features of large language models (LLMs) is a central goal of interpretability. LLMs are commonly assumed to use superposition to repr

model-releasesarxiv-cs-lg
25 Jun 2026
Model Releases

Geometry-Aware Online Scheduling for LLM Serving: From Theoretical Bound to System Practice

DGX agent

arXiv:2606.22327v2 Announce Type: replace Abstract: The explosive demand for interactive Large Language Model serving has highlighted the management of the Key-Value cache's dynamic memory footprint a

model-releasesarxiv-cs-ai
25 Jun 2026
Research

Heterogeneous and Adept Snapshot Distillation for 3D Semantic Segmentation

DGX agent

arXiv:2606.25278v1 Announce Type: new Abstract: Multi-modal fusion and multi-model ensembling are prevalent in enhancing the performance of 3D semantic segmentation. Despite the impressive performance

researcharxiv-cs-cv
25 Jun 2026
Model Releases

LLM Performance on a Real, Double-Marked GCSE Benchmark

DGX agent

arXiv:2606.24973v1 Announce Type: new Abstract: We introduce a dataset of 32,534 double-marked real student responses to GCSE mock exams (GCSEs are the UK's national exams, taken at age ~16), spanning

model-releasesarxiv-cs-cl
25 Jun 2026
Model Releases

Operator Boosting Produces Pareto-Efficient PDE Surrogates

DGX agent

arXiv:2606.17460v2 Announce Type: replace Abstract: Neural operators are widely used as surrogate solution maps for partial differential equations (PDEs), but full-size models can be costly to store,

model-releasesarxiv-cs-lg
25 Jun 2026
Model Releases

SciRisk-Bench: A Risk-Dimension-Aware Benchmark for AI4Science Safety

DGX agent

arXiv:2606.18936v2 Announce Type: replace Abstract: Large language models (LLMs) are increasingly embedded in AI for Science (AI4Science) workflows, from scientific question answering and literature a

model-releasesarxiv-cs-ai
25 Jun 2026
Research

Transferable Attack against Face Swapping in an Extended Space

DGX agent

arXiv:2606.25376v1 Announce Type: new Abstract: Although deep Face Swapping (FS) models may benefit the entertainment industry, they pose severe threats to privacy and security. Existing protections,

researcharxiv-cs-cv
25 Jun 2026
Model Releases

Verifiable Manifest Signing and Transparency Enforcement for Secure MCP-Based LLM Pipelines

DGX agent

arXiv:2601.23132v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are increasingly deployed in tool-driven environments such as healthcare analytics, financial systems, retrieval-

model-releasesarxiv-cs-ai
25 Jun 2026
Model Releases

A Benchmark for Hallucination Detection in VLMs for Gastrointestinal Endoscopy

DGX agent

arXiv:2606.24115v1 Announce Type: cross Abstract: Vision-language models (VLMs) are prone to hallucination, which remains a major barrier to their safe deployment in clinical practice. To date, most h

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

AGORA: An Archive-Grounded Benchmark for Agentic Workplace Document Reasoning

DGX agent

arXiv:2606.24526v1 Announce Type: new Abstract: Large language models are increasingly deployed as agents that reason over documents rather than answer from parametric knowledge. We study archive-grou

model-releasesarxiv-cs-cl
24 Jun 2026
Model Releases

BioMedVR: Confusion-Aware Mixture-of-Prompt Experts for Biomedical Visual Reprogramming

DGX agent

arXiv:2606.24740v1 Announce Type: new Abstract: Recent advances in vision-language models (VLMs) such as CLIP have demonstrated strong generalization across natural-image domains. However, adapting th

model-releasesarxiv-cs-cv
24 Jun 2026
Model Releases

Dual-Branch Cross-Projection Debiasing through Diffusion-based Disentanglement

DGX agent

arXiv:2606.24161v1 Announce Type: new Abstract: Foundation models trained on biased datasets often rely on spurious correlations between target labels and non-causal attributes, resulting in poor gene

model-releasesarxiv-cs-cv
24 Jun 2026
Model Releases

Flow-Corrected Thompson Sampling for Non-Stationary Contextual Bandits

DGX agent

arXiv:2606.23933v1 Announce Type: cross Abstract: We study non-stationary linear contextual bandits where the reward model drifts over time, rendering classical contextual bandit algorithms brittle be

model-releasesarxiv-cs-lg
24 Jun 2026
Model Releases

Holistic Data Scheduler for LLM Pre-training via Multi-Objective Reinforcement Learning

DGX agent

arXiv:2606.24133v1 Announce Type: cross Abstract: The composition of training data, governed by the diversity of sources and their mixing strategy, is a cornerstone of Large Language Model (LLM) pre-t

model-releasesarxiv-cs-cl
24 Jun 2026
Applications

Lite Any Stereo V2: Faster and Stronger Efficient Zero-Shot Stereo Matching

DGX agent

arXiv:2606.24457v1 Announce Type: new Abstract: Recent advances in stereo matching have achieved remarkable accuracy, but often rely on large models, heavy computation, or additional foundation-model

applicationsarxiv-cs-cv
24 Jun 2026
Model Releases

Quantum ring all-reduce: communication and privacy advantages for distributed learning

DGX agent

arXiv:2606.20344v2 Announce Type: replace-cross Abstract: Machine learning models have scaled to unprecedented sizes, making training across distributed devices the de facto standard in the field. In

model-releasesarxiv-cs-lg
24 Jun 2026
Model Releases

Reasoning as Attractor Dynamics: Latent Memory Retrieval via Gibbs-Weighted Energy Minimization

DGX agent

arXiv:2606.24543v1 Announce Type: new Abstract: Large Language Models (LLMs) are traditionally viewed as autoregressive generators. However, from the perspective of collective computation, they functi

model-releasesarxiv-cs-lg
24 Jun 2026
Model Releases

Towards Spec Learning: Inference-Time Alignment from Preference Pairs

DGX agent

arXiv:2606.24004v1 Announce Type: cross Abstract: Steering a large language model (LLM) toward a desired behavior typically relies on an iterative process of hand-crafting a prompt based on a careful

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

When Helpfulness Overrides Causal Caution: Context-Dependent Suppression and Recovery in LLMs

DGX agent

arXiv:2606.24370v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly integrated into decision-support roles in business and policy contexts. While prior benchmark studies have

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

ZONOS2 Technical Report

DGX agent

arXiv:2606.24320v1 Announce Type: cross Abstract: We present ZONOS2 8B, our latest TTS model, which achieves state-of-the-art naturalness, prosody, and voice cloning fidelity. We improve upon Zonos-v0

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

A Verifiable Search Is Not a Learnable Chain-of-Thought

DGX agent

arXiv:2606.21884v1 Announce Type: new Abstract: It is tempting to assume any task solvable by a short program can be taught to a model as its chain-of-thought: write the steps out, fine-tune, and the

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

ASCII Art Turns LLMs into VLA Controllers

DGX agent

arXiv:2606.21470v1 Announce Type: cross Abstract: Vision--Language--Action (VLA) controllers are often built by extending vision--language models (VLMs) with action supervision, relying on multimodal

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Beyond Fixed Budgets: Characterizing the Inelasticity and Limitations of Tree-of-Thought Reasoning Strategies

DGX agent

arXiv:2606.20599v1 Announce Type: cross Abstract: Tree of Thought (ToT) search has become a promising direction for improving the reasoning capabilities of large language models, but deploying these m

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Beyond 'One Language, One Script': Quantifying Orthographic Bias in Multilingual VLMs with PuMVR

DGX agent

arXiv:2606.20770v1 Announce Type: cross Abstract: Current Vision-Language Models (VLMs) are celebrated for their multilingual capabilities, yet they operate under a flawed assumption: that one languag

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Chehre: An Emoji-Prompted Video Dataset for Perceptually Diverse Facial Expression Recognition

DGX agent

arXiv:2606.21657v1 Announce Type: new Abstract: Facial expressions are nonverbal social signals used in human interaction, but facial expression recognition datasets often focus on static images, basi

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Chem2Gen-Bench: Benchmarking Chemical-to-Genetic Translation in Perturbation Response Space

DGX agent

arXiv:2606.21109v1 Announce Type: new Abstract: Virtual-cell and perturbation models are increasingly used to predict cellular responses for biomedical discovery, but chemical and genetic perturbation

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Confidently Wrong: Severity-Aware Calibration of Prompt-Injection Detectors under Attack Shift

DGX agent

arXiv:2606.22659v1 Announce Type: cross Abstract: Prompt-injection detectors are deployed as guards: a model scores an input and a downstream system trusts or blocks it on that score. I study the conf

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Double-Diffusion: Balancing Speed, Accuracy, and Uncertainty in Probabilistic Forecasting for Urban Sensor Networks

DGX agent

arXiv:2506.23053v3 Announce Type: replace Abstract: Urban sensor networks need forecasts that are accurate, carry useful uncertainty, and refresh fast enough to act on as new readings arrive. These go

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

ELDiff: When Evidential Learning Meets Text-to-Image Diffusion

DGX agent

arXiv:2606.20924v1 Announce Type: new Abstract: In multi-object text-to-image (T2I) diffusion, ensuring semantic consistency between textual prompts and generated visual content is crucial for image s

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Fara-1.5: Scalable Learning Environments for Computer Use Agents

DGX agent

arXiv:2606.20785v1 Announce Type: cross Abstract: Collecting computer use data from human demonstrations is expensive and slow, motivating the need for scalable generation strategies. This requires tw

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Good-Enough LLM Obfuscation (GELO)

DGX agent

arXiv:2603.05035v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are increasingly served on shared accelerators where an adversary with read access to device memory can observe K

model-releasesarxiv-cs-lg
23 Jun 2026
Safety

Happy Young Women, Grumpy Old Men? Emotion-Driven Demographic Biases in Synthetic Face Generation

DGX agent

arXiv:2602.00032v3 Announce Type: replace-cross Abstract: Synthetic faces from text-to-image (T2I) models pervade digital media, yet their demographic biases under emotionally conditioned prompts rema

safetyarxiv-cs-cv
23 Jun 2026
Tutorials

Keep The Essentials: Efficient Reference Conditioned Generation via Token Dropping

DGX agent

arXiv:2606.23682v1 Announce Type: new Abstract: Reference-based diffusion models enable highly controllable image generation by leveraging elements from input images to guide prompt-driven synthesis.

tutorialsarxiv-cs-cv
23 Jun 2026
Model Releases

LAYUP: Asynchronous decentralized gradient descent with LAYer-wise UPdates

DGX agent

arXiv:2410.05985v4 Announce Type: replace Abstract: The increasing size of deep learning models has made distributed training across multiple devices essential. Synchronous, centralized methods incur

model-releasesarxiv-cs-lg
23 Jun 2026
Local Ai

Local Causal Attribution of Chain-of-Thought Reasoning

DGX agent

arXiv:2606.21821v1 Announce Type: new Abstract: Understanding the causal structure of a language model's thought process is a problem of significant importance for both transparency and safety. In thi

local-aiarxiv-cs-lg
23 Jun 2026
Model Releases

MMGist: A Comprehensive Multimodal Benchmark for 2027

DGX agent

arXiv:2606.22437v1 Announce Type: new Abstract: We conduct a systematic study of 18 widely used vision-language benchmarks and identify three major issues: 1) many items do not rely on visual cues and

model-releasesarxiv-cs-cv
23 Jun 2026
← Previous
1…313314315316317…1065
Next →