AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
85,136Total entries
1Added by human
85,135Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,778 results
Model Releases

Continual Model Routing in Evolving Model Hubs

DGX agent

arXiv:2605.28577v1 Announce Type: new Abstract: AI model hubs provide access to a rapidly growing collection of powerful pre-trained models, enabling off-the-shelf mixture-of-experts systems with diff

model-releasesarxiv-cs-ai
28 May 2026
Model Releases
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

ConvMemory: A Lightweight Learned Memory Reranker, a Negative Attribution Result, and a Research-Preview Conflict Editor

DGX agent

arXiv:2605.28062v1 Announce Type: new Abstract: We describe ConvMemory, a small 3.6M-parameter learned reranker for conversational long-term memory retrieval, trained with cross-encoder teacher superv

model-releasesarxiv-cs-cl
28 May 2026
Model Releases

Cost-Sensitive Evaluation for Binary Classifiers

DGX agent

arXiv:2510.22016v2 Announce Type: replace Abstract: Selecting an appropriate evaluation metric for classifiers is crucial for model comparison, parameter optimization, and deployment decisions, yet th

model-releasesarxiv-cs-lg
28 May 2026
Model Releases

Cultural Binding Heads in Language Models

DGX agent

arXiv:2605.28543v1 Announce Type: new Abstract: LLMs often default to equal treatment across cultural groups, even though context warrants differentiation: this is a lack of difference awareness. Usin

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Cultural Fidelity in English-to-Hindi Translation: A Preservation-Fluency Frontier for Gender Recoverability

DGX agent

arXiv:2605.27654v1 Announce Type: cross Abstract: Generative translation systems are cultural technologies because they decide how socially meaningful cues are rendered within culturally specific gram

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

CyberJurors: A Multi-Agent Simulation Task for E-Commerce Disputes Verdict

DGX agent

arXiv:2605.28369v1 Announce Type: new Abstract: E-commerce platforms have begun recruiting crowdsourced jurors to adjudicate massive volumes of transaction disputes. Unlike formal legal judgment, E-co

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Dark Quest II: A Wide-Coverage Neural Network Emulator of the Nonlinear Matter Power Spectrum Across Extended Cosmologies

DGX agent

arXiv:2605.28596v1 Announce Type: cross Abstract: extsc{DarkEmulator2} is a neural network emulator of the nonlinear matter power spectrum in a nine-dimensional w_0 w_a nu o CDM parameter space, devel

model-releasesarxiv-cs-lg
28 May 2026
Model Releases

Data-Efficient On-Policy Distillation for Automatic Speech Recognition

DGX agent

arXiv:2605.28139v1 Announce Type: new Abstract: Building competitive automatic speech recognition (ASR) models usually requires large-scale au- dio supervision, which makes reproduction and specializa

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Debate with Images: Detecting Deceptive Behaviors in Multimodal Large Language Models

DGX agent

arXiv:2512.00349v2 Announce Type: replace Abstract: Are frontier AI systems becoming more capable? Certainly. Yet such progress is not an unalloyed blessing but rather a Trojan horse: behind their per

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Decentralized Parameter-Free Online Learning with Compressed Gossip

DGX agent

arXiv:2605.27831v1 Announce Type: new Abstract: We study decentralized online convex optimization when agents communicate over a graph and messages may be compressed. Classical decentralized online me

model-releasesarxiv-cs-lg
28 May 2026
Model Releases

DecomposeRL: Learning to Ask Useful, Informative, and Diverse Questions for Semi-Supervised, Traceable Claim Verification

DGX agent

arXiv:2605.27858v1 Announce Type: cross Abstract: Claim verification splits between end-to-end classifiers that are accurate but yields no inspectable traces, and decomposition-based methods produce i

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Decoupled Training with Local Reinforcement Fine-Tuning in Federated Learning

DGX agent

arXiv:2605.27900v1 Announce Type: new Abstract: Federated Learning (FL) with pre-trained Vision-Language Models (VLMs) has emerged as a promising paradigm for various downstream tasks. By leveraging i

model-releasesarxiv-cs-cv
28 May 2026
Model Releases

Deepfake-Eval-2024: A Multi-Modal In-the-Wild Benchmark of Deepfakes Circulated in 2024

DGX agent

arXiv:2503.02857v5 Announce Type: replace-cross Abstract: In the age of increasingly realistic generative AI, robust deepfake detection is essential for mitigating fraud and disinformation. While many

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

DeepSciVerify: Verifying Scientific Claim--Citation Alignment via LLM-Driven Evidence Escalation

DGX agent

arXiv:2605.27710v1 Announce Type: new Abstract: Misalignment between claims and their cited evidence is a common failure mode in reports generated by large language models, limiting their reliability

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Deformable Gaussian Occupancy: Decoupling Rigid and Nonrigid Motion with Factorized Distillation

DGX agent

arXiv:2605.28587v1 Announce Type: new Abstract: Understanding dynamic 3D environments is essential for safe autonomous driving, particularly when reasoning about human-centric, nonrigid agents. Howeve

model-releasesarxiv-cs-cv
28 May 2026
Model Releases

DEPART: DEcomposing PARiTy across Multilingual LLMs

DGX agent

arXiv:2605.28163v1 Announce Type: cross Abstract: Multilingual Large Language Models (mLLMs) leaderboards report per-language accuracy but rarely explain why disparities emerge, leaving systemic biase

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Detection Without Correction: A Two-Parameter Decomposition of Multi-Stage LLM Pipelines

DGX agent

arXiv:2605.27559v1 Announce Type: cross Abstract: Multi-stage LLM pipelines that perform multi-agent debate, intrinsic self-correction, or retrieval-augmented verification exhibit puzzling aggregate b

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

'Developers can update Claude’s instructions mid-task without breaking the prompt cache or routing the update through a user turn' wtf? how?…

DGX agent

'Developers can update Claude’s instructions mid-task without breaking the prompt cache or routing the update through a user turn' wtf? how?? Introducing Claude Opus 4.8: it builds on Opus 4.7 with sh

model-releasesswyx--x
28 May 2026
Model Releases

Did this actually happen? It seems very suspicious.

DGX agent

Did this actually happen? It seems very suspicious. 'An AI consultant tells Axios one of their clients recently spent half a billion dollars in a single month after failing to put usage limits on Clau

model-releasesethan-mollick--x
28 May 2026
Model Releases

Differential syntactic and semantic encoding in LLMs

DGX agent

arXiv:2601.04765v4 Announce Type: replace-cross Abstract: We study how syntactic and semantic information is encoded in inner layer representations of Large Language Models (LLMs), focusing on the ver

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Diffusion-Based Ukrainian Handwritten Text Generation with Cross-Domain Style Transfer

DGX agent

arXiv:2605.27487v1 Announce Type: cross Abstract: Handwritten text generation (HTG) conditioned on writer style has been widely studied for Latin scripts, but remains underexplored for low-resource an

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Dimensionality Reduction for Robust Federated Learning: A Theoretical Analysis and Convergence Guarantee

DGX agent

arXiv:2605.28335v1 Announce Type: new Abstract: Federated Learning (FL) enables multiple clients to collaboratively train models without sharing raw data, but it is highly vulnerable to Byzantine atta

model-releasesarxiv-cs-lg
28 May 2026
Model Releases

DisasterBench: Benchmarking LLM Planning under Typed Tool Interface Constraints

DGX agent

arXiv:2605.27957v1 Announce Type: new Abstract: Disasters cause severe societal impacts, demanding rapid coordination of heterogeneous AI tools, from satellite analysis to flood prediction and damage

model-releasesarxiv-cs-cl
28 May 2026
Model Releases

Disentangling Language Roles in Multilingual LLM Task Execution

DGX agent

arXiv:2605.27649v1 Announce Type: new Abstract: Multilingual LLMs are increasingly used when instruction, source content, and required response languages do not coincide. Existing benchmarks have expa

model-releasesarxiv-cs-cl
28 May 2026
Model Releases

Do Agents Think Deeper? A Mechanistic Investigation of Layer-Wise Dynamics in Sequential Planning

DGX agent

arXiv:2605.27935v1 Announce Type: new Abstract: Recent mechanistic studies suggest that large language models (LLMs) may utilize their depth inefficiently in standard single-turn tasks. Whether this s

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Do Audio LLMs Listen or Read? Analyzing and Mitigating Paralinguistic Failures with VoxParadox

DGX agent

arXiv:2605.27772v1 Announce Type: cross Abstract: Audio large language models (Audio LLMs) demonstrate strong performance on speech understanding tasks, yet their ability to understand paralinguistic

model-releasesarxiv-cs-lg
28 May 2026
Model Releases

Do Clinical Models Change Treatment Decisions?

DGX agent

arXiv:2605.28129v1 Announce Type: new Abstract: Clinical foundation models are evaluated with factual or exam-style medical QA, but treatment decisions must change when patient context changes. We int

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Do LLMs Build World Models From Text? A Multilingual Diagnostic of Spatial Reasoning

DGX agent

arXiv:2605.28277v1 Announce Type: new Abstract: Whether large language models (LLMs) construct internal spatial world models from pure-text descriptions remains contested, and whether such capabilitie

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Do LLMs Favor Their Providers? Measuring Vertical Integration Bias in Code Generation

DGX agent

arXiv:2605.28515v1 Announce Type: cross Abstract: Large Language Models (LLMs) have become an integral part of software development, especially with the advent of agentic capabilities. Yet, many front

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Do Models Know Why They Changed Their Mind? Interpretability and Faithfulness of Chain-of-Thought Under Knowledge Conflict

DGX agent

arXiv:2605.27773v1 Announce Type: cross Abstract: When a language model sees a document contradicting its training knowledge, it must choose: follow the document or trust itself. Prior work proved thi

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Do We Really Need Quantum Machine Learning?: A Multidimensional Empirical Study

DGX agent

arXiv:2605.27923v1 Announce Type: cross Abstract: The rapid growth of computer vision and increasingly complex image recognition tasks has exposed fundamental computational limitations of classical ma

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Dr-CiK: A Testbed for Foresight-Driven Agents

DGX agent

arXiv:2605.27904v1 Announce Type: new Abstract: Time series forecasting in real-world settings often depends not only on historical observations, but also on external context that must be actively dis

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

DriveWAM: Video Generative Priors Enable Scalable World-Action Modeling for Autonomous Driving

DGX agent

arXiv:2605.28544v1 Announce Type: new Abstract: Pretrained foundation models have become an important basis for end-to-end autonomous driving. In contrast to vision-language models pretrained primaril

model-releasesarxiv-cs-cv
28 May 2026
Model Releases

DRTriton: Large-Scale Synthetic Data Driven Reinforcement Learning for Triton Kernel Generation

DGX agent

arXiv:2603.21465v2 Announce Type: replace Abstract: Developing efficient CUDA kernels is a fundamental yet challenging task in the generative AI industry. Recent research leverages Large Language Mode

model-releasesarxiv-cs-cl
28 May 2026
Model Releases

DynaSchedBench: Calibrated Dynamic Scheduling Benchmarks and Observability Paradox in LLM-based Scheduling Agents

DGX agent

arXiv:2605.27566v1 Announce Type: new Abstract: Progress in neural combinatorial optimization for Dynamic Flexible Job Shop Scheduling Problem (DFJSP) is currently hindered by a methodological tension

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Efficient and Scalable Provenance Tracking for LLM-Generated Code Snippets

DGX agent

arXiv:2605.28510v1 Announce Type: cross Abstract: Large language models (LLMs) for code completion and generation are increasingly used in software development, yet they may reproduce training example

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Efficient Pre-Training of LLMs through Truncated SVD Layers

DGX agent

arXiv:2605.28573v1 Announce Type: cross Abstract: The massive scaling of Large Language Models (LLMs) has made pretraining increasingly cost-prohibitive. While low-rank representation and orthonormal

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

EgoBench: An Interactive Egocentric Multimodal Benchmark for Tool-Using Agents

DGX agent

arXiv:2605.27820v1 Announce Type: new Abstract: As AI agents increasingly operate in open, real-world environments, they require a deep synergy of multimodal perception, tool invocation with multi-hop

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Energy-Structured Low-Rank Adaptation for Continual Learning

DGX agent

arXiv:2605.27482v1 Announce Type: cross Abstract: While orthogonal subspace methods try to mitigate task interference in Continual Learning (CL), they often suffer from energy diffusion across the bas

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Enhancing Trustworthy GUI Grounding via Self-Critiqued Reinforcement Learning

DGX agent

arXiv:2510.27266v2 Announce Type: replace Abstract: Autonomous graphical user interface (GUI) agents rely on accurate GUI grounding, which maps language instructions to on-screen coordinates, to execu

model-releasesarxiv-cs-cv
28 May 2026
Model Releases

ESC-Skills: Discovering and Self-Evolving Skills for Emotional Support Conversations

DGX agent

arXiv:2605.27908v1 Announce Type: cross Abstract: Existing emotional support conversation (ESC) systems mainly rely on end-to-end response generation or coarse strategy supervision, offering limited i

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

EVADE-Bench: Multimodal Benchmark for Evaluating and Enhancing Evasive Content Detection

DGX agent

arXiv:2505.17654v4 Announce Type: replace-cross Abstract: E-commerce platforms increasingly rely on Large Language Models (LLMs) and Vision Language Models (VLMs) to detect illicit or misleading produ

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Evaluating Local Explainability Metrics for Machine Learning Models on Tabular Data

DGX agent

arXiv:2605.27618v1 Announce Type: new Abstract: Despite the wide use of explainability techniques to attempt to understand the behavior of Artificial Intelligence (AI), the generated explanations may

model-releasesarxiv-cs-lg
28 May 2026
Model Releases

EventShiftFlow: Towards Hardware-efficient FPGA-based Flow Estimation

DGX agent

arXiv:2605.28312v1 Announce Type: cross Abstract: Event-based vision sensors offer asynchronous, high-temporal-resolution measurements that are attractive for low-latency robotic perception, but many

model-releasesarxiv-cs-cv
28 May 2026
Model Releases

Every9D-21M: Large-Scale Real-World 9D Canonicalization of Everyday Objects

DGX agent

arXiv:2605.28270v1 Announce Type: new Abstract: Estimating the 9D pose of everyday objects from a single real-world image remains challenging. This is largely due to the lack of large-scale supervisio

model-releasesarxiv-cs-cv
28 May 2026
Model Releases

Evolving Dataflow to process massive datasets for machine learning

DGX agent

Google created MapReduce more than 20 years ago to solve the scaling problems in data processing that the then young company was running into. The AI era that we are in now demands efficient, large-sc

model-releasesgoogle-cloud-ai
28 May 2026
Model Releases

EvoSpec: Evolving Speculative Decoding via Real-Time Vocabulary and Parameter AdaptationTarget

DGX agent

arXiv:2605.27390v1 Announce Type: cross Abstract: Speculative decoding accelerates Large Language Model inference via a draft-then-verify paradigm, yet the output projection layer becomes a bottleneck

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Exploratory Experience Shapes the Geometry of Predictive Representations

DGX agent

arXiv:2605.27929v1 Announce Type: cross Abstract: Active sensing links behavior and learning through an action-perception loop: actions determine the observations used to update internal predictive mo

model-releasesarxiv-cs-lg
28 May 2026
← Previous
1…253254255256257…475
Next →