AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,288 results
Model Releases

CalBrief: A Pilot Diagnostic Benchmark for Evidence-Calibrated Scientific Briefing with Large Language Models

DGX agent

arXiv:2606.27383v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used as research assistants, yet it remains unclear whether they can calibrate research takeaways to the

model-releasesarxiv-cs-ai
29 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Causal Connections: Leveraging Multilingual Fine-Tuning for Financial QA@FinCausal 2026

DGX agent

arXiv:2606.27446v1 Announce Type: new Abstract: This paper describes team HSA_CORAL's submission to the FinCausal 2026 shared task on extracting cause-effect relations from financial narratives via ex

model-releasesarxiv-cs-cl
29 Jun 2026
Model Releases

CausalFlip: A Benchmark for LLM Causal Judgment Beyond Semantic Matching

DGX agent

arXiv:2602.20094v2 Announce Type: replace Abstract: As large language models (LLMs) witness increasing deployment in complex, high-stakes decision-making scenarios, it becomes imperative to ground the

model-releasesarxiv-cs-ai
29 Jun 2026
Model Releases

Complex-Valued 2D Gaussian Representation for Computer-Generated Holography

DGX agent

arXiv:2511.15022v2 Announce Type: replace Abstract: Complex-valued Gaussian primitives have recently been explored for representing holographic radiance fields in 3D novel view synthesis. In this work

model-releasesarxiv-cs-cv
29 Jun 2026
Model Releases

Compression-Driven Anomaly Detection in Brain MRI Using an Interpretable Quantum Autoencoder

DGX agent

arXiv:2606.27411v1 Announce Type: cross Abstract: We study a quantum autoencoder (QAE) for compression-driven anomaly detection in brain MRI data. The approach leverages angle encoding to map image pa

model-releasesarxiv-cs-ai
29 Jun 2026
Model Releases

Contagion Networks: Evaluator Preference Propagation in Multi-Agent LLM Systems

DGX agent

arXiv:2606.20493v2 Announce Type: replace-cross Abstract: When large language models serve as evaluators in multi-agent systems, their strategy preferences -- whether induced by explicit prompts or by

model-releasesarxiv-cs-ai
29 Jun 2026
Model Releases

Context-specific Credibility-aware Multimodal Fusion with Conditional Probabilistic Circuits

DGX agent

arXiv:2603.26629v2 Announce Type: replace Abstract: Multimodal fusion requires integrating information from multiple sources that may conflict depending on context. Existing fusion approaches typicall

model-releasesarxiv-cs-lg
29 Jun 2026
Model Releases

Curriculum-guided Change Detection Training: Toward Accurate Serac Fall Monitoring

DGX agent

arXiv:2606.28012v1 Announce Type: new Abstract: Change Detection (CD) aims to identify semantic or structural changes from nearly registered multi-temporal images. While recent advances in training me

model-releasesarxiv-cs-cv
29 Jun 2026
Model Releases

Deepfake Media Generation and Detection in the Generative AI Era: A Survey and Outlook

DGX agent

arXiv:2411.19537v2 Announce Type: replace-cross Abstract: We survey deepfake generation and detection techniques, covering all deepfake media types: image, video, audio and multimodal content. We iden

model-releasesarxiv-cs-ai
29 Jun 2026
Model Releases

DMind Benchmark: Toward a Holistic Assessment of LLM Capabilities across the Web3 Domain

DGX agent

arXiv:2504.16116v4 Announce Type: replace-cross Abstract: The Web3 ecosystem, underpinned by cryptographic primitives and decentralized consensus, represents a high-stakes environment where software v

model-releasesarxiv-cs-ai
29 Jun 2026
Model Releases

DMV-Bench: Diagnosing Long-Horizon Multimodal Agents' Visual Memory with Incidental Cue Injection

DGX agent

arXiv:2606.27499v1 Announce Type: cross Abstract: Research on agent memory has matured rapidly, but almost entirely on the text side: few existing benchmarks ask, in an interactive environment, when a

model-releasesarxiv-cs-ai
29 Jun 2026
Model Releases

Do Speech Emphasis Models Generalize across Languages and Emotions?

DGX agent

arXiv:2606.27717v1 Announce Type: cross Abstract: Prosodic emphasis varies across languages, emotions, and speaking styles, yet existing emphasis detection models are largely trained and evaluated on

model-releasesarxiv-cs-ai
29 Jun 2026
Model Releases

End-to-End Dynamic Sparsity for Resource-Adaptive LLM Inference

DGX agent

arXiv:2606.27743v1 Announce Type: cross Abstract: Large Language Models (LLMs) inference is typically deployed under a static resource assumption, where models execute a fixed computational graph rega

model-releasesarxiv-cs-ai
29 Jun 2026
Model Releases

Estimation--Prediction Tradeoff in Causal Probabilistic Temporal Graphs

DGX agent

arXiv:2606.28225v1 Announce Type: new Abstract: Temporal link prediction is usually evaluated by predictive performance on unseen edges, but in probabilistic temporal graphs this criterion can conflat

model-releasesarxiv-cs-lg
29 Jun 2026
Model Releases

EXPLORE-Bench: Egocentric Scene Prediction with Long-Horizon Reasoning

DGX agent

arXiv:2603.09731v3 Announce Type: replace-cross Abstract: Multimodal large language models (MLLMs) are increasingly considered as a foundation for embodied agents, yet it remains unclear whether they

model-releasesarxiv-cs-ai
29 Jun 2026
Model Releases

Fine-tuning a multimodal large language model for clinician-grade autism behavioral scoring from short home videos

DGX agent

arXiv:2606.27484v1 Announce Type: new Abstract: Autism spectrum disorder (ASD) affects 1 in 31 US children, yet median age at diagnosis exceeds four years. Artificial intelligence pipelines that provi

model-releasesarxiv-cs-cv
29 Jun 2026
Model Releases

FoggyTrust: Robust Federated Learning with Hierarchical Trust Networks

DGX agent

arXiv:2606.27622v1 Announce Type: new Abstract: Byzantine-robust federated learning seeks to protect distributed model training from malicious or corrupted clients without requiring access to their pr

model-releasesarxiv-cs-lg
29 Jun 2026
Model Releases

Formalizing Latent Thoughts: Four Axioms of Thought Representation in LLMs

DGX agent

arXiv:2606.27378v1 Announce Type: new Abstract: We introduce an axiomatic evaluation framework for latent thought representations in LLMs, comprising metrics that are independent of downstream benchma

model-releasesarxiv-cs-cl
29 Jun 2026
Model Releases

Freshness and the Limits of Heuristic Trend Detection in Temporal RAG

DGX agent

arXiv:2509.19376v2 Announce Type: replace-cross Abstract: We present a lightweight, model-agnostic temporal layer for RAG and use cybersecurity data to separate two problems that are usually conflated

model-releasesarxiv-cs-ai
29 Jun 2026
Model Releases

From Black-Box to Clinical Insight: A Multi-Stage Explainable Framework for Speech-Based Cognitive Impairment Detection

DGX agent

arXiv:2606.27973v1 Announce Type: cross Abstract: Speech-based cognitive impairment detection offers a noninvasive, accessible alternative to costly biomarker assays, yet transformer-based models rema

model-releasesarxiv-cs-ai
29 Jun 2026
Model Releases

From Detection to Action: Using LLM Agents for Fault-Tolerant Control

DGX agent

arXiv:2606.28011v1 Announce Type: cross Abstract: We propose an agentic Large Language Model (LLM) framework for active Fault-Tolerant Control (FTC) that transforms fault detection outputs into constr

model-releasesarxiv-cs-lg
29 Jun 2026
Model Releases

From Signals to Transfer: A Factorised Study of Probe-Based Uncertainty Estimation in Large Language Models

DGX agent

arXiv:2606.27679v1 Announce Type: cross Abstract: Probe-based uncertainty estimation (UE) has emerged as a prominent approach to detect hallucinations in Large Language Models (LLMs) by learning uncer

model-releasesarxiv-cs-ai
29 Jun 2026
Model Releases

GAIA: A Data Flywheel System for Training GUI Test-Time Scaling Critic Models

DGX agent

arXiv:2601.18197v2 Announce Type: replace Abstract: While Large Vision-Language Models (LVLMs) have significantly advanced GUI agents' capabilities in parsing textual instructions, interpreting screen

model-releasesarxiv-cs-ai
29 Jun 2026
Model Releases

Govern the Repository, Not the Agent: Measuring Ecosystem-Level Risk in AI-Native Software

DGX agent

arXiv:2606.28235v1 Announce Type: cross Abstract: Autonomous coding agents now open and merge pull requests in shared repositories at scale, and the field evaluates them the way it has always evaluate

model-releasesarxiv-cs-ai
29 Jun 2026
Model Releases

GRAFT: Biological Graph and Hypergraph Benchmarks for Linked Gene Expression and Phenotypic Trait Prediction in Arabidopsis thaliana

DGX agent

arXiv:2606.27413v1 Announce Type: cross Abstract: Understanding which genes control which traits in an organism remains one of the central challenges in biology. Despite significant advances in data c

model-releasesarxiv-cs-ai
29 Jun 2026
Model Releases

Graph Dimensionality Reduction for Contextual Bandits: Structure-Specific Regret Bounds under Approximate Smoothness and Noisy Eigenspaces

DGX agent

arXiv:2606.27917v1 Announce Type: new Abstract: Contextual bandits with graph-structured arms arise in recommendation, citation retrieval, and social advertising, where arms connected on a graph tend

model-releasesarxiv-cs-lg
29 Jun 2026
Model Releases

HAT-4D: Lifting Monocular Video for 4D Multi-Object Interactions via Human-Agent Collaboration

DGX agent

arXiv:2606.28215v1 Announce Type: cross Abstract: Extracting dynamic 4D object interactions from massive, in-the-wild monocular videos offers a highly efficient data collection pathway for scaling Emb

model-releasesarxiv-cs-ai
29 Jun 2026
Model Releases

Health-ORSC-Bench: A Benchmark for Measuring Over-Refusal and Safety Completion in Health Context

DGX agent

arXiv:2601.17642v2 Announce Type: replace Abstract: Safety alignment in Large Language Models is critical for healthcare; however, reliance on binary refusal boundaries often results in over-refusal o

model-releasesarxiv-cs-ai
29 Jun 2026
Model Releases

HumanMoveVQA: Can Video MLLMs reason about human movement in videos?

DGX agent

arXiv:2606.27999v1 Announce Type: new Abstract: Despite the rapid advance of Multimodal Large Language Models (MLLMs) in high-level video understanding, a fundamental bottleneck remains: these models

model-releasesarxiv-cs-cv
29 Jun 2026
Model Releases

Hybrid Fact-Checking that Integrates Knowledge Graphs, Large Language Models, and Search-Based Retrieval Agents Improves Interpretable Claim Verification

DGX agent

arXiv:2511.03217v2 Announce Type: replace-cross Abstract: Large language models (LLMs) excel in generating fluent utterances but can lack reliable grounding in verified information. At the same time,

model-releasesarxiv-cs-ai
29 Jun 2026
Model Releases

Hyperellipsoid Density Sampling: Exploitative Sequences to Accelerate High-Dimensional Optimization

DGX agent

arXiv:2511.07836v4 Announce Type: replace-cross Abstract: The curse of dimensionality remains a persistent challenge in modern optimization problems. Expanding the search space into higher dimensions

model-releasesarxiv-cs-lg
29 Jun 2026
Model Releases

iCost: A Novel Instance-Complexity-Based Cost-Sensitive Learning Framework

DGX agent

arXiv:2409.13007v3 Announce Type: replace-cross Abstract: Class imbalance poses a significant challenge in classification tasks, often causing standard learning algorithms to become biased toward the

model-releasesarxiv-cs-ai
29 Jun 2026
Model Releases

Ko-WideSearch: A Korean Breadth-Search Benchmark for Exhaustive Set Enumeration by Web Agents

DGX agent

arXiv:2606.27595v1 Announce Type: new Abstract: Web-agent benchmarks overwhelmingly measure depth -- pinning one obscure answer behind a chain of constraints -- while breadth, exhaustively enumerating

model-releasesarxiv-cs-cl
29 Jun 2026
Model Releases

Learning in Markovian bandits with non-observable states and constrained decision epochs

DGX agent

arXiv:2606.27448v1 Announce Type: new Abstract: This paper studies the problem of regret minimization in Markovian bandits with non-observable states and possibly constrained decision epochs. The focu

model-releasesarxiv-cs-lg
29 Jun 2026
Model Releases

Learning Peer Influence Probabilities with Linear Contextual Bandits

DGX agent

arXiv:2510.19119v2 Announce Type: replace Abstract: In networked environments, it is common for users to share recommendations about content, products, services, and possible courses of action. Whethe

model-releasesarxiv-cs-lg
29 Jun 2026
Model Releases

Learning to Evict from Key-Value Cache

DGX agent

arXiv:2602.10238v2 Announce Type: replace Abstract: The growing size of Large Language Models (LLMs) makes efficient inference challenging, primarily due to the memory demands of the autoregressive Ke

model-releasesarxiv-cs-cl
29 Jun 2026
Model Releases

LLawCo: Learning Laws of Cooperation for Modeling Embodied Multi-Agent Behavior

DGX agent

arXiv:2606.28182v1 Announce Type: cross Abstract: Embodied agents operating in decentralized and partially observable environments have attracted growing attention in recent years. However, existing l

model-releasesarxiv-cs-ai
29 Jun 2026
Model Releases

LocalNav: Distilling Frontier VLMs and Embodied RL for On-Device Object Goal Navigation

DGX agent

arXiv:2606.27871v1 Announce Type: new Abstract: Vision Language Models (VLMs) have emerged in the robotic domain as a powerful tool that enables environmental perception with language context, serving

model-releasesarxiv-cs-ro
29 Jun 2026
Model Releases

Lost at the End: Primacy Bias in Multimodal Retrieval-Augmented Question Answering

DGX agent

arXiv:2606.16494v2 Announce Type: replace-cross Abstract: Knowledge-based visual question answering (KB-VQA) lets vision-language systems answer questions that exceed their parametric knowledge by con

model-releasesarxiv-cs-ai
29 Jun 2026
Model Releases

MemoBench: Benchmarking World Modeling in Dynamically Changing Environments

DGX agent

arXiv:2606.27537v1 Announce Type: new Abstract: Video generation models aspire to simulate dynamic environments, and several benchmarks now evaluate memory consistency across frames. However, most ass

model-releasesarxiv-cs-cv
29 Jun 2026
Model Releases

Mitigating LLM-based p-Hacking by Preregistering for the Next LLM

DGX agent

arXiv:2606.27687v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used to generate, classify, and annotate data whose outputs feed downstream hypothesis tests. However, L

model-releasesarxiv-cs-ai
29 Jun 2026
Model Releases

MLVC: Multi-platform Learned Video Codec for Real-World Deployment

DGX agent

arXiv:2606.28027v1 Announce Type: cross Abstract: Neural video codecs have surpassed classical codecs in coding efficiency but remain impractical for deployment due to cross-platform incompatibility a

model-releasesarxiv-cs-ai
29 Jun 2026
Model Releases

MobileManiBench: Simplifying Model Verification for Mobile Manipulation

DGX agent

arXiv:2602.05233v2 Announce Type: replace Abstract: Vision-language-action models have advanced robotic manipulation but remain constrained by reliance on the large, teleoperation-collected datasets d

model-releasesarxiv-cs-ro
29 Jun 2026
Model Releases

Monocular Avatar Reconstruction via Cascaded Diffusion Priors and UV-Space Differentiable Shading

DGX agent

arXiv:2606.28144v1 Announce Type: new Abstract: Reconstructing high-fidelity, relightable 3D avatars from a single in-the-wild image is a challenging ill-posed problem, primarily hindered by the scarc

model-releasesarxiv-cs-cv
29 Jun 2026
Model Releases

Monte Carlo with kernel-based Gibbs measures: Guarantees for probabilistic herding

DGX agent

arXiv:2402.11736v3 Announce Type: replace Abstract: Kernel herding belongs to a family of deterministic quadratures that seek to minimize the maximum mean discrepancy (MMD), that is, the worst-case in

model-releasesarxiv-cs-lg
29 Jun 2026
Model Releases

Mosaic: A Benchmark Suite for Differentiable Physics Solvers

DGX agent

arXiv:2606.27895v1 Announce Type: cross Abstract: Differentiable partial differential equation (PDE) solvers underpin solver-in-the-loop ML training, gradient-based optimal control, and inverse proble

model-releasesarxiv-cs-lg
29 Jun 2026
Model Releases

MultiHashFormer: Hash-based Generative Language Models

DGX agent

arXiv:2606.28057v1 Announce Type: cross Abstract: Language models (LMs) represent tokens using embedding matrices that scale linearly with the vocabulary size. To constrain the parameter footprint, pr

model-releasesarxiv-cs-ai
29 Jun 2026
Model Releases

Multimodal Evaluator Preference Collapse: Cross-Modal Coupling in Self-Evolving Agents

DGX agent

arXiv:2606.16682v3 Announce Type: replace-cross Abstract: When AI agents use language models to evaluate their own outputs in a feedback loop, systematic biases emerge. We show that Evaluator Preferen

model-releasesarxiv-cs-cl
29 Jun 2026
← Previous
1…111112113114115…361
Next →