AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,419
  • Agents7,559
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,170
  • Local Ai4,934
  • Model Releases23,909
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,419
  • Agents7,559
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,170
  • Local Ai4,934
  • Model Releases23,909
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlog
88,419Total entries
1Added by human
88,418Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,638 results
Model Releases

From Black-Box to Clinical Insight: A Multi-Stage Explainable Framework for Speech-Based Cognitive Impairment Detection

DGX agent

arXiv:2606.27973v1 Announce Type: cross Abstract: Speech-based cognitive impairment detection offers a noninvasive, accessible alternative to costly biomarker assays, yet transformer-based models rema

model-releasesarxiv-cs-ai
29 Jun 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Safety

HunyuanImage 3.0 Technical Report

DGX agent

arXiv:2509.23951v3 Announce Type: replace Abstract: We present HunyuanImage 3.0, a native multimodal model that unifies multimodal understanding and generation within an autoregressive framework, with

safetyarxiv-cs-cv
29 Jun 2026
Model Releases

LocalNav: Distilling Frontier VLMs and Embodied RL for On-Device Object Goal Navigation

DGX agent

arXiv:2606.27871v1 Announce Type: new Abstract: Vision Language Models (VLMs) have emerged in the robotic domain as a powerful tool that enables environmental perception with language context, serving

model-releasesarxiv-cs-ro
29 Jun 2026
Model Releases

Multimodal Evaluator Preference Collapse: Cross-Modal Coupling in Self-Evolving Agents

DGX agent

arXiv:2606.16682v3 Announce Type: replace-cross Abstract: When AI agents use language models to evaluate their own outputs in a feedback loop, systematic biases emerge. We show that Evaluator Preferen

model-releasesarxiv-cs-cl
29 Jun 2026
Model Releases

NormAct: A Benchmark for Hidden Social Norm Compliance in Embodied Planning

DGX agent

arXiv:2606.27826v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) are increasingly deployed as embodied planners in egocentric environments, where task success requires not only

model-releasesarxiv-cs-ai
29 Jun 2026
Applications

Optimizing Teacher-Student Partitioning for Scalable Knowledge Distillation on HPC Systems

DGX agent

arXiv:2606.27797v1 Announce Type: cross Abstract: Knowledge Distillation (KD) enables training smaller student models under the guidance of larger teacher models, and the widely adopted TRL library im

applicationsarxiv-cs-ai
29 Jun 2026
Model Releases

Speculative Refinement: A Hybrid Autoregressive Diffusion Decoding Strategy and Its Behavior Across Benchmarks

DGX agent

arXiv:2606.27474v1 Announce Type: cross Abstract: How should we evaluate generation systems that combine autoregressive (AR) and diffusion decoding? We study this question through Speculative Refineme

model-releasesarxiv-cs-ai
29 Jun 2026
Tutorials

How to keep AI spend flat while token usage grows exponentially: Not with friction and spend alerts. With better defaults, routing, and cach…

DGX agent

How to keep AI spend flat while token usage grows exponentially: Not with friction and spend alerts. With better defaults, routing, and caching. Better Defaults (not Usage Caps) – Engineers can choose

tutorialsclem-delangue--x
27 Jun 2026
Model Releases

Assert, don't describe: Linguistic features that shift LLM reasoning about animal welfare

DGX agent

arXiv:2606.26104v1 Announce Type: cross Abstract: Animal-welfare advocates produce a lot of writing, and increasingly that writing trains the language models that millions of people then ask about ani

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

Autoregressive Boltzmann Generators

DGX agent

arXiv:2606.27361v1 Announce Type: cross Abstract: Efficient sampling of molecular systems at thermodynamic equilibrium is a hallmark challenge in statistical physics. This challenge has driven the dev

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

Cascaded Multi-Granularity Pruning for On-Device LLM Inference in Industrial IoT

DGX agent

arXiv:2606.26861v1 Announce Type: new Abstract: Deploying large language models (LLMs) on Industrial Internet of Things (IIoT) edge devices demands extreme compression, yet existing structured pruning

model-releasesarxiv-cs-cl
26 Jun 2026
Model Releases

Context Recycling for Long-Horizon LLM Inference

DGX agent

arXiv:2606.26105v1 Announce Type: cross Abstract: Large language models (LLMs) exhibit strong capabilities in short-context reasoning but degrade in performance over long conversational horizons due t

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

CORTEX: A Structured Reasoning Benchmark for Trustworthy 3D Chest CT MLLMs

DGX agent

arXiv:2606.27264v1 Announce Type: new Abstract: Reasoning in multimodal large language models (MLLMs) has shown strong promise in medical imaging. However, this reasoning is usually free-form text jud

model-releasesarxiv-cs-cv
26 Jun 2026
Model Releases

DiCoBench: Benchmarking Multi-Image Fine-Grained Perception via Differential and Commonality Visual Cues

DGX agent

arXiv:2606.26602v1 Announce Type: new Abstract: Recent advancements in Multimodal Large Language Models (MLLMs) have demonstrated impressive fine-grained perception capabilities. However, existing ben

model-releasesarxiv-cs-cv
26 Jun 2026
Model Releases

Empirical Software Engineering TerraProbe: A Layered-Oracle Framework for Detecting Deceptive Fixes in LLM-Assisted Terraform

DGX agent

arXiv:2606.26590v1 Announce Type: new Abstract: Security misconfigurations in Terraform Infrastructure-as-Code are a growing risk in cloud deployments, and large language models are increasingly used

model-releasesarxiv-cs-lg
26 Jun 2026
Model Releases

From Guessing to Placeholding: A Cost-Theoretic Framework for Uncertainty-Aware Code Completion

DGX agent

arXiv:2604.01849v2 Announce Type: replace Abstract: While Large Language Models (LLMs) have demonstrated exceptional proficiency in code completion, they typically adhere to a Hard Completion (HC) par

model-releasesarxiv-cs-cl
26 Jun 2026
Model Releases

Helpfulness Hurts: Domain-Dependent Degradation of Mid-Trained Compassion Values Under Post-Training

DGX agent

arXiv:2606.26102v1 Announce Type: cross Abstract: Standard post-training pipelines apply supervised fine-tuning (SFT) and reinforcement learning (RL) to make language models helpful, but these process

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

How Do Tool-Augmented LLM Agents Perform on Real-World Energy Analytics Tasks?

DGX agent

arXiv:2606.26346v1 Announce Type: new Abstract: Agentic benchmarks have emerged across general-purpose and domain-specific settings, including finance, coding, law, and drug discovery, yet energy-doma

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

Implementation of reinforcement learning in chemical reaction networks: application to phototaxis as curiosity-driven exploration

DGX agent

arXiv:2606.26168v1 Announce Type: new Abstract: Living systems navigate environments using noisy and incomplete sensory signals. In unicellular algae, phototaxis is often modeled as a mechanistic run-

model-releasesarxiv-cs-lg
26 Jun 2026
Model Releases

Mean-Field PhiBE: Continuous-Time Mean-Field Reinforcement Learning from Discrete-Time Data

DGX agent

arXiv:2606.26498v1 Announce Type: cross Abstract: This paper addresses model-free continuous-time mean-field control in a setting where the population dynamics evolve continuously according to an unkn

model-releasesarxiv-cs-lg
26 Jun 2026
Model Releases

OI-Bench: An Option Injection Benchmark for Evaluating LLM Susceptibility to Directive Interference

DGX agent

arXiv:2601.13300v2 Announce Type: replace Abstract: Benchmarking large language models (LLMs) is critical for understanding their capabilities, limitations, and robustness. In addition to interface ar

model-releasesarxiv-cs-cl
26 Jun 2026
Model Releases

OpenAI introduces GPT-5.6 to challenge Claude Mythos 5

DGX agent

OpenAI Group PBC today introduced GPT-5.6, a new series of large language models that it says can outperform Claude Mythos 5 across certain coding tasks. The most advanced algorithm in the lineup is k

model-releasessiliconangle
26 Jun 2026
Model Releases

Perception, Verdict, and Evolution: Hindsight-Driven Self-Refining Forensics Agent for AI-Generated Image Detection

DGX agent

arXiv:2606.26552v1 Announce Type: cross Abstract: The rapid advancement of generative models presents a significant challenge to existing deepfake detection methods, particularly given the widespread

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

PhyEditBench: A Real-World Multi-Stage Benchmark for Physics-Aware Image Editing

DGX agent

arXiv:2606.26551v1 Announce Type: new Abstract: While instruction-based image editing, enabled by multi-modal generative models, has advanced significantly, existing benchmarks lack a comprehensive ev

model-releasesarxiv-cs-cv
26 Jun 2026
Model Releases

Spurious Rewards Paradox: Mechanistically Understanding How RLVR Activates Memorization Shortcuts in LLMs

DGX agent

arXiv:2601.11061v2 Announce Type: replace-cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) is highly effective for enhancing LLM reasoning, yet recent evidence shows models like Q

model-releasesarxiv-cs-cl
26 Jun 2026
Model Releases

What We are Missing in Multimodal LLM Evaluation?

DGX agent

arXiv:2606.26348v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) can process diverse inputs, e.g., text, images, audio, and video, and generate textual responses. While their c

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

2K Retrofit: Entropy-Guided Efficient Sparse Refinement for High-Resolution 3D Geometry Prediction

DGX agent

arXiv:2603.19964v3 Announce Type: replace Abstract: High-resolution geometric prediction is essential for robust perception in autonomous driving, robotics, and AR/MR, but current foundation models ar

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

BOFA: Bridge-Layer Orthogonal Low-Rank Fusion for CLIP-Based Class-Incremental Learning

DGX agent

arXiv:2511.11421v2 Announce Type: replace Abstract: Class-Incremental Learning (CIL) aims to continually learn new categories without forgetting previously acquired knowledge. Vision-language models s

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

C3-Bench: A Context-Aware Change Captioning Benchmark

DGX agent

arXiv:2606.25445v1 Announce Type: new Abstract: While Change Captioning systems have garnered substantial attention to respond to our evolving world, their true performance on diverse real-world chang

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

CausalRAG2: Hierarchical Causal Knowledge Graph Design for RAG

DGX agent

arXiv:2602.05143v2 Announce Type: replace Abstract: Retrieval augmented generation (RAG) has enhanced large language models by enabling access to external knowledge, with graph-based RAG emerging as a

model-releasesarxiv-cs-ai
25 Jun 2026
Model Releases

Curvature-Guided Mixing for MLLM Adaptation

DGX agent

arXiv:2606.24963v1 Announce Type: new Abstract: Fine-tuning Multimodal Large Language Models (MLLMs) on specialized tasks often leads to catastrophic forgetting of their general capabilities. Existing

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

Dream at SemEval-2026 Task 13: SALSA for Single-Pass Machine-Generated Code Detection

DGX agent

arXiv:2606.25102v1 Announce Type: new Abstract: Large language models have transformed code generation, raising concerns around authorship, assessment integrity, and software trust. SemEval-2026 Task

model-releasesarxiv-cs-cl
25 Jun 2026
Model Releases

Evidence for feature-specific error correction in LLMs

DGX agent

arXiv:2606.24964v1 Announce Type: new Abstract: Understanding the features of large language models (LLMs) is a central goal of interpretability. LLMs are commonly assumed to use superposition to repr

model-releasesarxiv-cs-lg
25 Jun 2026
Model Releases

Geometry-Aware Online Scheduling for LLM Serving: From Theoretical Bound to System Practice

DGX agent

arXiv:2606.22327v2 Announce Type: replace Abstract: The explosive demand for interactive Large Language Model serving has highlighted the management of the Key-Value cache's dynamic memory footprint a

model-releasesarxiv-cs-ai
25 Jun 2026
Research

Heterogeneous and Adept Snapshot Distillation for 3D Semantic Segmentation

DGX agent

arXiv:2606.25278v1 Announce Type: new Abstract: Multi-modal fusion and multi-model ensembling are prevalent in enhancing the performance of 3D semantic segmentation. Despite the impressive performance

researcharxiv-cs-cv
25 Jun 2026
Model Releases

LLM Performance on a Real, Double-Marked GCSE Benchmark

DGX agent

arXiv:2606.24973v1 Announce Type: new Abstract: We introduce a dataset of 32,534 double-marked real student responses to GCSE mock exams (GCSEs are the UK's national exams, taken at age ~16), spanning

model-releasesarxiv-cs-cl
25 Jun 2026
Model Releases

Operator Boosting Produces Pareto-Efficient PDE Surrogates

DGX agent

arXiv:2606.17460v2 Announce Type: replace Abstract: Neural operators are widely used as surrogate solution maps for partial differential equations (PDEs), but full-size models can be costly to store,

model-releasesarxiv-cs-lg
25 Jun 2026
Model Releases

SciRisk-Bench: A Risk-Dimension-Aware Benchmark for AI4Science Safety

DGX agent

arXiv:2606.18936v2 Announce Type: replace Abstract: Large language models (LLMs) are increasingly embedded in AI for Science (AI4Science) workflows, from scientific question answering and literature a

model-releasesarxiv-cs-ai
25 Jun 2026
Research

Transferable Attack against Face Swapping in an Extended Space

DGX agent

arXiv:2606.25376v1 Announce Type: new Abstract: Although deep Face Swapping (FS) models may benefit the entertainment industry, they pose severe threats to privacy and security. Existing protections,

researcharxiv-cs-cv
25 Jun 2026
Model Releases

Verifiable Manifest Signing and Transparency Enforcement for Secure MCP-Based LLM Pipelines

DGX agent

arXiv:2601.23132v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are increasingly deployed in tool-driven environments such as healthcare analytics, financial systems, retrieval-

model-releasesarxiv-cs-ai
25 Jun 2026
Model Releases

A Benchmark for Hallucination Detection in VLMs for Gastrointestinal Endoscopy

DGX agent

arXiv:2606.24115v1 Announce Type: cross Abstract: Vision-language models (VLMs) are prone to hallucination, which remains a major barrier to their safe deployment in clinical practice. To date, most h

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

AGORA: An Archive-Grounded Benchmark for Agentic Workplace Document Reasoning

DGX agent

arXiv:2606.24526v1 Announce Type: new Abstract: Large language models are increasingly deployed as agents that reason over documents rather than answer from parametric knowledge. We study archive-grou

model-releasesarxiv-cs-cl
24 Jun 2026
Model Releases

BioMedVR: Confusion-Aware Mixture-of-Prompt Experts for Biomedical Visual Reprogramming

DGX agent

arXiv:2606.24740v1 Announce Type: new Abstract: Recent advances in vision-language models (VLMs) such as CLIP have demonstrated strong generalization across natural-image domains. However, adapting th

model-releasesarxiv-cs-cv
24 Jun 2026
Model Releases

Dual-Branch Cross-Projection Debiasing through Diffusion-based Disentanglement

DGX agent

arXiv:2606.24161v1 Announce Type: new Abstract: Foundation models trained on biased datasets often rely on spurious correlations between target labels and non-causal attributes, resulting in poor gene

model-releasesarxiv-cs-cv
24 Jun 2026
Model Releases

Flow-Corrected Thompson Sampling for Non-Stationary Contextual Bandits

DGX agent

arXiv:2606.23933v1 Announce Type: cross Abstract: We study non-stationary linear contextual bandits where the reward model drifts over time, rendering classical contextual bandit algorithms brittle be

model-releasesarxiv-cs-lg
24 Jun 2026
Model Releases

Holistic Data Scheduler for LLM Pre-training via Multi-Objective Reinforcement Learning

DGX agent

arXiv:2606.24133v1 Announce Type: cross Abstract: The composition of training data, governed by the diversity of sources and their mixing strategy, is a cornerstone of Large Language Model (LLM) pre-t

model-releasesarxiv-cs-cl
24 Jun 2026
Applications

Lite Any Stereo V2: Faster and Stronger Efficient Zero-Shot Stereo Matching

DGX agent

arXiv:2606.24457v1 Announce Type: new Abstract: Recent advances in stereo matching have achieved remarkable accuracy, but often rely on large models, heavy computation, or additional foundation-model

applicationsarxiv-cs-cv
24 Jun 2026
Model Releases

Quantum ring all-reduce: communication and privacy advantages for distributed learning

DGX agent

arXiv:2606.20344v2 Announce Type: replace-cross Abstract: Machine learning models have scaled to unprecedented sizes, making training across distributed devices the de facto standard in the field. In

model-releasesarxiv-cs-lg
24 Jun 2026
← Previous
1…390391392393394…1326
Next →