AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,630Total entries
1Added by human
84,629Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,288 results
Model Releases

Can Segmentation Models Understand the World? Towards Proactive Affordance Reasoning via Visual Chain-of-Thought

DGX agent

arXiv:2605.27764v1 Announce Type: cross Abstract: Recent segmentation models couple large language models (LLMs) with mask decoders to ground complex language expressions into masks, yet their instruc

model-releasesarxiv-cs-ai
28 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

CAREF: Calibration-Aware Regularization for Explanation Faithfulness Without Rationale Supervision

DGX agent

arXiv:2605.27835v1 Announce Type: cross Abstract: We introduce CAREF, a parameter-efficient fine-tuning framework that jointly optimizes predictive accuracy and explanation faithfulness via calibratio

model-releasesarxiv-cs-cl
28 May 2026
Model Releases

Category-Level 3D Correspondence in Camera Space via Morphable Object Priors

DGX agent

arXiv:2605.28257v1 Announce Type: new Abstract: Understanding 3D objects from images is fundamental to robotics and AR/VR applications. While recent work has made progress in category-level pose estim

model-releasesarxiv-cs-cv
28 May 2026
Model Releases

CFDTwin: An open-source GUI and Python toolkit for POD-NN surrogate modeling of ANSYS Fluent simulations

DGX agent

arXiv:2605.27725v1 Announce Type: cross Abstract: High-fidelity computational fluid dynamics (CFD) is widely used for thermal-fluid design, but repeated CFD solves remain expensive for design optimiza

model-releasesarxiv-cs-lg
28 May 2026
Model Releases

ChildEval: When large language models meet children's personalities

DGX agent

arXiv:2605.27805v1 Announce Type: cross Abstract: While LLMs enable personalized chatbots, their effectiveness in child-centered personalization remains unclear, as systematic evaluation of child-spec

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Chinese Word Boundary Recovery through Character Alignment Projection

DGX agent

arXiv:2605.28128v1 Announce Type: new Abstract: Chinese word segmentation is especially fragile in non-standard text, where language learner errors and other character-level divergences disrupt the wo

model-releasesarxiv-cs-cl
28 May 2026
Model Releases

Chirpy3D: Part-Aware Multi-View Diffusion for Creative Fine-Grained Object Generation

DGX agent

arXiv:2501.04144v3 Announce Type: replace Abstract: Understanding and generating the fine-grained structure of objects -- such as birds with species-specific beaks, wings, and tails -- is a long-stand

model-releasesarxiv-cs-cv
28 May 2026
Model Releases

CiteCheck: Retrieval-Grounded Detection of LLM Citation Hallucinations in Scientific Text

DGX agent

arXiv:2605.27700v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used to generate scientific reports, but they can produce references that appear plausible while contain

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

ClinConsensus: A Physician-Calibrated Benchmark for Evaluating Clinical Rubric Coverage in Chinese Medical LLMs

DGX agent

arXiv:2603.02097v5 Announce Type: replace Abstract: Open-ended medical LLM evaluation remains weakly grounded in physician-calibrated coverage of clinically relevant response criteria, especially in l

model-releasesarxiv-cs-cl
28 May 2026
Model Releases

ClinicalAgents: Multi-Agent Orchestration for Clinical Decision Making with Dual-Memory

DGX agent

arXiv:2603.26182v2 Announce Type: replace Abstract: While Large Language Models (LLMs) have demonstrated potential in healthcare, they often struggle with the complex, non-linear reasoning required fo

model-releasesarxiv-cs-cl
28 May 2026
Model Releases

Code as a Weapon: A Consensus-Labeled Prompt Bank for Measuring Coding-Model Compliance with Malicious-Code Requests

DGX agent

arXiv:2605.28734v1 Announce Type: cross Abstract: A general-purpose language model that answers a harmful question returns text; a coding model that complies with a malicious request can return a work

model-releasesarxiv-cs-cl
28 May 2026
Model Releases

CogPortrait: Fine-Grained Eye-Region Control in Portrait Animation via Hierarchical Agent Planning

DGX agent

arXiv:2605.28056v1 Announce Type: new Abstract: Portrait animation methods have achieved substantial visual quality and lip synchronization, but fine-grained manipulation of the eye region still faces

model-releasesarxiv-cs-cv
28 May 2026
Model Releases

CogVLA: Cognition-Aligned Vision-Language-Action Model via Instruction-Driven Routing & Sparsification

DGX agent

arXiv:2508.21046v3 Announce Type: replace Abstract: Recent Vision-Language-Action (VLA) models built on pre-trained Vision-Language Models (VLMs) require extensive post-training, resulting in high com

model-releasesarxiv-cs-cv
28 May 2026
Model Releases

Coherence Collapse: Diagnosing Why Code Agents Fail After Reaching the Right Code

DGX agent

arXiv:2603.24631v2 Announce Type: replace-cross Abstract: Code agents resolve 65-70% of SWE-bench Verified issues, but Pass@1 cannot tell us why the rest fail, and, as we show, capable-model failures

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Colosseum V2: Benchmarking Generalization for Vision Language Action Models

DGX agent

arXiv:2605.27759v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models demonstrate promising generalization in robotic manipulation, driven by advances in large-scale vision and language

model-releasesarxiv-cs-ro
28 May 2026
Model Releases

Comparative Analysis of Liquid Neural Networks and LSTM for Sequential Pattern Recognition: Robustness, Efficiency, and Clinical Utility

DGX agent

arXiv:2605.27467v1 Announce Type: cross Abstract: Traditional Recurrent Neural Networks (RNNs) and Long Short-Term Memory (LSTM) units operate on discrete time steps, often failing to capture the flui

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Confidence-Orchestrated Self-Evolution against Uncertain LLM Feedback

DGX agent

arXiv:2605.28010v1 Announce Type: new Abstract: Self-evolving large language models (LLMs) learn by generating their own training tasks and solutions, reducing reliance on human-curated supervision. H

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

ConRAG: Consensus-Driven Multi-View Retrieval for Multi-Hop Question Answering

DGX agent

arXiv:2605.28093v1 Announce Type: new Abstract: Retrieval-augmented generation (RAG) has emerged as a promising paradigm for enhancing large language models (LLMs) on multi-hop question answering (QA)

model-releasesarxiv-cs-cl
28 May 2026
Model Releases

Conservative neural posterior estimation via distributionally robust training

DGX agent

arXiv:2605.28516v1 Announce Type: cross Abstract: Simulation-based inference with neural posterior estimation (NPE) often yields overconfident and unreliable posteriors under limited simulation budget

model-releasesarxiv-cs-lg
28 May 2026
Model Releases

Continual Model Routing in Evolving Model Hubs

DGX agent

arXiv:2605.28577v1 Announce Type: new Abstract: AI model hubs provide access to a rapidly growing collection of powerful pre-trained models, enabling off-the-shelf mixture-of-experts systems with diff

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

ConvMemory: A Lightweight Learned Memory Reranker, a Negative Attribution Result, and a Research-Preview Conflict Editor

DGX agent

arXiv:2605.28062v1 Announce Type: new Abstract: We describe ConvMemory, a small 3.6M-parameter learned reranker for conversational long-term memory retrieval, trained with cross-encoder teacher superv

model-releasesarxiv-cs-cl
28 May 2026
Model Releases

Cost-Sensitive Evaluation for Binary Classifiers

DGX agent

arXiv:2510.22016v2 Announce Type: replace Abstract: Selecting an appropriate evaluation metric for classifiers is crucial for model comparison, parameter optimization, and deployment decisions, yet th

model-releasesarxiv-cs-lg
28 May 2026
Model Releases

Cultural Binding Heads in Language Models

DGX agent

arXiv:2605.28543v1 Announce Type: new Abstract: LLMs often default to equal treatment across cultural groups, even though context warrants differentiation: this is a lack of difference awareness. Usin

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Cultural Fidelity in English-to-Hindi Translation: A Preservation-Fluency Frontier for Gender Recoverability

DGX agent

arXiv:2605.27654v1 Announce Type: cross Abstract: Generative translation systems are cultural technologies because they decide how socially meaningful cues are rendered within culturally specific gram

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

CyberJurors: A Multi-Agent Simulation Task for E-Commerce Disputes Verdict

DGX agent

arXiv:2605.28369v1 Announce Type: new Abstract: E-commerce platforms have begun recruiting crowdsourced jurors to adjudicate massive volumes of transaction disputes. Unlike formal legal judgment, E-co

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Dark Quest II: A Wide-Coverage Neural Network Emulator of the Nonlinear Matter Power Spectrum Across Extended Cosmologies

DGX agent

arXiv:2605.28596v1 Announce Type: cross Abstract: extsc{DarkEmulator2} is a neural network emulator of the nonlinear matter power spectrum in a nine-dimensional w_0 w_a nu o CDM parameter space, devel

model-releasesarxiv-cs-lg
28 May 2026
Model Releases

Data-Efficient On-Policy Distillation for Automatic Speech Recognition

DGX agent

arXiv:2605.28139v1 Announce Type: new Abstract: Building competitive automatic speech recognition (ASR) models usually requires large-scale au- dio supervision, which makes reproduction and specializa

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Debate with Images: Detecting Deceptive Behaviors in Multimodal Large Language Models

DGX agent

arXiv:2512.00349v2 Announce Type: replace Abstract: Are frontier AI systems becoming more capable? Certainly. Yet such progress is not an unalloyed blessing but rather a Trojan horse: behind their per

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Decentralized Parameter-Free Online Learning with Compressed Gossip

DGX agent

arXiv:2605.27831v1 Announce Type: new Abstract: We study decentralized online convex optimization when agents communicate over a graph and messages may be compressed. Classical decentralized online me

model-releasesarxiv-cs-lg
28 May 2026
Model Releases

DecomposeRL: Learning to Ask Useful, Informative, and Diverse Questions for Semi-Supervised, Traceable Claim Verification

DGX agent

arXiv:2605.27858v1 Announce Type: cross Abstract: Claim verification splits between end-to-end classifiers that are accurate but yields no inspectable traces, and decomposition-based methods produce i

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Decoupled Training with Local Reinforcement Fine-Tuning in Federated Learning

DGX agent

arXiv:2605.27900v1 Announce Type: new Abstract: Federated Learning (FL) with pre-trained Vision-Language Models (VLMs) has emerged as a promising paradigm for various downstream tasks. By leveraging i

model-releasesarxiv-cs-cv
28 May 2026
Model Releases

Deepfake-Eval-2024: A Multi-Modal In-the-Wild Benchmark of Deepfakes Circulated in 2024

DGX agent

arXiv:2503.02857v5 Announce Type: replace-cross Abstract: In the age of increasingly realistic generative AI, robust deepfake detection is essential for mitigating fraud and disinformation. While many

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

DeepSciVerify: Verifying Scientific Claim--Citation Alignment via LLM-Driven Evidence Escalation

DGX agent

arXiv:2605.27710v1 Announce Type: new Abstract: Misalignment between claims and their cited evidence is a common failure mode in reports generated by large language models, limiting their reliability

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Deformable Gaussian Occupancy: Decoupling Rigid and Nonrigid Motion with Factorized Distillation

DGX agent

arXiv:2605.28587v1 Announce Type: new Abstract: Understanding dynamic 3D environments is essential for safe autonomous driving, particularly when reasoning about human-centric, nonrigid agents. Howeve

model-releasesarxiv-cs-cv
28 May 2026
Model Releases

DEPART: DEcomposing PARiTy across Multilingual LLMs

DGX agent

arXiv:2605.28163v1 Announce Type: cross Abstract: Multilingual Large Language Models (mLLMs) leaderboards report per-language accuracy but rarely explain why disparities emerge, leaving systemic biase

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Detection Without Correction: A Two-Parameter Decomposition of Multi-Stage LLM Pipelines

DGX agent

arXiv:2605.27559v1 Announce Type: cross Abstract: Multi-stage LLM pipelines that perform multi-agent debate, intrinsic self-correction, or retrieval-augmented verification exhibit puzzling aggregate b

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Differential syntactic and semantic encoding in LLMs

DGX agent

arXiv:2601.04765v4 Announce Type: replace-cross Abstract: We study how syntactic and semantic information is encoded in inner layer representations of Large Language Models (LLMs), focusing on the ver

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Diffusion-Based Ukrainian Handwritten Text Generation with Cross-Domain Style Transfer

DGX agent

arXiv:2605.27487v1 Announce Type: cross Abstract: Handwritten text generation (HTG) conditioned on writer style has been widely studied for Latin scripts, but remains underexplored for low-resource an

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Dimensionality Reduction for Robust Federated Learning: A Theoretical Analysis and Convergence Guarantee

DGX agent

arXiv:2605.28335v1 Announce Type: new Abstract: Federated Learning (FL) enables multiple clients to collaboratively train models without sharing raw data, but it is highly vulnerable to Byzantine atta

model-releasesarxiv-cs-lg
28 May 2026
Model Releases

DisasterBench: Benchmarking LLM Planning under Typed Tool Interface Constraints

DGX agent

arXiv:2605.27957v1 Announce Type: new Abstract: Disasters cause severe societal impacts, demanding rapid coordination of heterogeneous AI tools, from satellite analysis to flood prediction and damage

model-releasesarxiv-cs-cl
28 May 2026
Model Releases

Disentangling Language Roles in Multilingual LLM Task Execution

DGX agent

arXiv:2605.27649v1 Announce Type: new Abstract: Multilingual LLMs are increasingly used when instruction, source content, and required response languages do not coincide. Existing benchmarks have expa

model-releasesarxiv-cs-cl
28 May 2026
Model Releases

Do Agents Think Deeper? A Mechanistic Investigation of Layer-Wise Dynamics in Sequential Planning

DGX agent

arXiv:2605.27935v1 Announce Type: new Abstract: Recent mechanistic studies suggest that large language models (LLMs) may utilize their depth inefficiently in standard single-turn tasks. Whether this s

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Do Audio LLMs Listen or Read? Analyzing and Mitigating Paralinguistic Failures with VoxParadox

DGX agent

arXiv:2605.27772v1 Announce Type: cross Abstract: Audio large language models (Audio LLMs) demonstrate strong performance on speech understanding tasks, yet their ability to understand paralinguistic

model-releasesarxiv-cs-lg
28 May 2026
Model Releases

Do Clinical Models Change Treatment Decisions?

DGX agent

arXiv:2605.28129v1 Announce Type: new Abstract: Clinical foundation models are evaluated with factual or exam-style medical QA, but treatment decisions must change when patient context changes. We int

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Do LLMs Build World Models From Text? A Multilingual Diagnostic of Spatial Reasoning

DGX agent

arXiv:2605.28277v1 Announce Type: new Abstract: Whether large language models (LLMs) construct internal spatial world models from pure-text descriptions remains contested, and whether such capabilitie

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Do LLMs Favor Their Providers? Measuring Vertical Integration Bias in Code Generation

DGX agent

arXiv:2605.28515v1 Announce Type: cross Abstract: Large Language Models (LLMs) have become an integral part of software development, especially with the advent of agentic capabilities. Yet, many front

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Do Models Know Why They Changed Their Mind? Interpretability and Faithfulness of Chain-of-Thought Under Knowledge Conflict

DGX agent

arXiv:2605.27773v1 Announce Type: cross Abstract: When a language model sees a document contradicting its training knowledge, it must choose: follow the document or trust itself. Prior work proved thi

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Do We Really Need Quantum Machine Learning?: A Multidimensional Empirical Study

DGX agent

arXiv:2605.27923v1 Announce Type: cross Abstract: The rapid growth of computer vision and increasingly complex image recognition tasks has exposed fundamental computational limitations of classical ma

model-releasesarxiv-cs-ai
28 May 2026
← Previous
1…186187188189190…361
Next →