AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,630Total entries
1Added by human
84,629Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,288 results
Model Releases

Evi-Steer: Learning to Steer Biomedical Vision-Language Models through Efficient and Generalizable Evidential Tuning

DGX agent

arXiv:2605.26292v1 Announce Type: cross Abstract: Parameter-efficient adaptation of vision-language foundation models is crucial for precise multimodal understanding of biomedical images, yet existing

model-releasesarxiv-cs-cl
27 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Exploiting Local Dynamics Regularity for Reusable Skills in Offline Hierarchical RL

DGX agent

arXiv:2605.26371v1 Announce Type: new Abstract: Hierarchical Reinforcement Learning (HRL) promises to solve long-horizon Reinforcement Learning (RL) tasks more efficiently than non-hierarchical counte

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

Extra-Merge: Tracing the Rank-1 Subspace of Model Merging in Language Model Pre-Training

DGX agent

arXiv:2605.26484v1 Announce Type: new Abstract: Model merging has emerged as a lightweight paradigm for enhancing Large Language Models (LLMs), yet its underlying mechanisms remain poorly understood.

model-releasesarxiv-cs-lg
27 May 2026
Model Releases

FAB-Bench: A Framework for Adaptive RAG Benchmarking in Semiconductor Manufacturing

DGX agent

arXiv:2605.26476v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) has become critical for knowledge-intensive applications, yet evaluating its performance in vertical domains remain

model-releasesarxiv-cs-cl
27 May 2026
Model Releases

Faithfulness Evaluation for Decoder-only LLM Attributions with Controlled Retained Information

DGX agent

arXiv:2601.03089v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are increasingly evaluated with input attribution methods, yet comparing such explanations remains challenging. E

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

Falcon-X: A Time Series Foundation Model for Heterogeneous Multivariate Modeling

DGX agent

arXiv:2605.27286v1 Announce Type: cross Abstract: Time series foundation models (TSFMs) are transforming the forecasting paradigm through large-scale cross-domain pretraining. However, most existing T

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

FedTreeLoRA: Reconciling Statistical and Functional Heterogeneity in Federated LoRA Fine-Tuning

DGX agent

arXiv:2603.13282v2 Announce Type: replace-cross Abstract: Federated Learning (FL) with Low-Rank Adaptation (LoRA) has become a standard for privacy-preserving LLM fine-tuning. However, existing person

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

FineVLA: Fine-Grained Instruction Alignment for Steerable Vision-Language-Action Policies

DGX agent

arXiv:2605.27284v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models are increasingly expected to not only complete robot tasks, but also follow human instructions about how those tas

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

Focal Reward: Balanced Reinforcement Learning under Rubric-Based Rewards

DGX agent

arXiv:2605.26579v1 Announce Type: new Abstract: The open-ended generation in LLMs usually requires multi-dimensional rubrics to adequately assess quality and guide the improvement of reinforcement lea

model-releasesarxiv-cs-lg
27 May 2026
Model Releases

From PDF to RAG-Ready: Evaluating Document Conversion Frameworks for Domain-Specific Question Answering

DGX agent

arXiv:2604.04948v2 Announce Type: replace-cross Abstract: Retrieval-Augmented Generation (RAG) systems depend critically on the quality of document preprocessing, yet no prior study has evaluated PDF

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

GEM: Geometric Entropy Mixing for Optimal LLM Data Curation

DGX agent

arXiv:2605.26121v1 Announce Type: cross Abstract: LLM pre-training efficacy increasingly depends on data composition rather than sheer volume. Yet, optimal mixing is hindered by categorization flaws:

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

Gemini Embedding 2: A Native Multimodal Embedding Model from Gemini

DGX agent

arXiv:2605.27295v1 Announce Type: new Abstract: We introduce Gemini Embedding 2, a native multimodal embedding model that allows embedding video, audio, image, and text modalities in a unified represe

model-releasesarxiv-cs-cv
27 May 2026
Model Releases

GeoFaith: A Spatio-Temporal Dual View of Faithful Chain-of-Thought

DGX agent

arXiv:2605.26893v1 Announce Type: cross Abstract: Chain-of-Thought (CoT) reasoning has advanced large language models (LLMs), but outcome-based supervision leads to pervasive post-hoc rationalization,

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

Geometry-Aware Representation Denoising for Robust Multi-view 3D Reconstruction

DGX agent

arXiv:2605.26230v1 Announce Type: new Abstract: Multi-view 3D reconstruction has achieved remarkable progress with the advent of feed-forward 3D reconstruction models. However, these models are typica

model-releasesarxiv-cs-cv
27 May 2026
Model Releases

'Give Me BF16 or Give Me Death'? Accuracy-Performance Trade-Offs in LLM Quantization

DGX agent

arXiv:2411.02355v4 Announce Type: replace-cross Abstract: Quantization is a powerful tool for accelerating large language model (LLM) inference, but the accuracy-performance trade-offs across differen

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

Grammar of the Wave: Towards Explainable Multivariate Time Series Event Detection via Neuro-Symbolic VLM Agents

DGX agent

arXiv:2603.11479v2 Announce Type: replace-cross Abstract: Time Series Event Detection (TSED) aims to localize semantically meaningful events in time series data, with critical applications in high-sta

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

GraphDancer: Training LLMs to Explore and Reason over Graphs via Two-Stage Curriculum Post-Training

DGX agent

arXiv:2602.02518v2 Announce Type: replace-cross Abstract: Large language models (LLMs) increasingly rely on external knowledge to improve factuality, yet many real-world knowledge sources are organize

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

Guiding Token-Sparse Diffusion Models

DGX agent

arXiv:2601.01608v2 Announce Type: replace Abstract: Diffusion models deliver high quality in image synthesis but remain expensive during training and inference. Recent works have leveraged the inheren

model-releasesarxiv-cs-cv
27 May 2026
Model Releases

Helicase: Uncertainty-Guided Supply Chain Knowledge Graph Construction with Autonomous Multi-Agent LLMs

DGX agent

arXiv:2605.26835v1 Announce Type: new Abstract: LLM-based multi-agent systems have been widely adopted for knowledge retrieval and report generation, synthesizing known information through web search

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

How Chain-of-Thought Works? Tracing Information Flow from Decoding, Projection, and Activation

DGX agent

arXiv:2507.20758v2 Announce Type: replace Abstract: Chain-of-Thought (CoT) prompting significantly enhances model reasoning, yet its internal mechanisms remain poorly understood. We analyze CoT's oper

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

HTMLCure: Turning Browser Experience into State Guided Repair for Interactive HTML

DGX agent

arXiv:2605.26807v1 Announce Type: cross Abstract: LLMs can now produce full HTML pages, but many of those pages are only superficially correct: they render once, then fail under scroll, hover, click,

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

Hubness, Not Anisotropy, Drives Cross-Lingual Retrieval Asymmetry in Multilingual Embedding Models

DGX agent

arXiv:2605.26575v1 Announce Type: new Abstract: Multilingual embedding models are deployed under the assumption that cross-lingual retrieval is symmetric: if a query in language A retrieves its transl

model-releasesarxiv-cs-cl
27 May 2026
Model Releases

ImViD: Immersive Volumetric Videos for Enhanced VR Engagement

DGX agent

arXiv:2503.14359v2 Announce Type: replace Abstract: User engagement is greatly enhanced by fully immersive multi-modal experiences that combine visual and auditory stimuli. Consequently, the next fron

model-releasesarxiv-cs-cv
27 May 2026
Model Releases

InfoQuant: Shaping Activation Distributions for Low-Bit LLM Quantization

DGX agent

arXiv:2605.26175v1 Announce Type: cross Abstract: Low-bit activation quantization remains a major bottleneck in efficient large language model (LLM) deployment. The difficulty is not only that activat

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

InfoSynth: Information-Guided Benchmark Synthesis for LLMs

DGX agent

arXiv:2601.00575v2 Announce Type: replace Abstract: Large language models (LLMs) have demonstrated significant advancements in reasoning and code generation, but efficiently creating new benchmarks to

model-releasesarxiv-cs-cl
27 May 2026
Model Releases

Intelligent Detection and Mitigation of Carpet-Bombing DDoS Attacks in SDN Using Retrieval-Augmented Generation and Large Language Models

DGX agent

arXiv:2605.26307v1 Announce Type: cross Abstract: Software-Defined Networking (SDN) provides flexible and programmable network management; however, its centralized control architecture remains highly

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

Interpretability and Generalization Bounds for Learning Spatial Physics

DGX agent

arXiv:2506.15199v3 Announce Type: replace Abstract: While there are many applications of ML to scientific problems that look promising, visuals can be deceiving. Using numerical analysis techniques, w

model-releasesarxiv-cs-lg
27 May 2026
Model Releases

InterSketch: An Interleaved Reasoning Model with Self-correcting Visual Sketch and Stepwise Reward

DGX agent

arXiv:2605.26520v1 Announce Type: cross Abstract: While vision-language models (VLMs) have exhibited multi-turn visual reasoning capabilities, their reasoning trajectories remain relatively shallow an

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

IPIBench: Evaluating Interactive Proactive Intelligence of MLLMs under Continuous Streams

DGX agent

arXiv:2605.27074v1 Announce Type: new Abstract: Recent multimodal large language models (MLLMs) achieve strong performance on reactive question answering, but real-world streaming assistants require p

model-releasesarxiv-cs-cv
27 May 2026
Model Releases

It's Not the Capability: Harness Sensitivity Is Non-Monotone Across LLM Agent Tiers

DGX agent

arXiv:2605.26731v1 Announce Type: new Abstract: A prevalent assumption in LLM agent deployment holds that more structured harnesses universally improve reliability, and that higher-capability models n

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

JobBench: Aligning Agent Work With Human Will

DGX agent

arXiv:2605.26329v1 Announce Type: new Abstract: Current benchmarks for occupational AI agents are scoped primarily by economic values, telling a replacement story. We introduce JobBench, which evaluat

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

JuICE: A Benchmark for Evaluating LLM-Judge in Identifying Cultural Errors

DGX agent

arXiv:2605.26955v1 Announce Type: cross Abstract: As large language models (LLMs) are increasingly deployed to users around the world, they are integrated into everyday tasks across diverse cultural c

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

Kandinsky 5.0: A Family of Foundation Models for Image and Video Generation

DGX agent

arXiv:2511.14993v3 Announce Type: replace-cross Abstract: This report introduces Kandinsky 5.0, a family of state-of-the-art foundation models for high-resolution image and 10-second video synthesis.

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

Knowledge Graphs as the Missing Data Layer for LLM-Based Industrial Asset Operations

DGX agent

arXiv:2605.26874v1 Announce Type: cross Abstract: LLM-based agents for industrial asset operations show limited accuracy when reasoning over flat document stores. AssetOpsBench (KDD 2026) establishes

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

L2Rec: Towards Dual-View Understanding of LLMs for Personalized Recommendation

DGX agent

arXiv:2605.26717v1 Announce Type: cross Abstract: Adapting large language models (LLMs) for personalized recommendation requires aligning their general-purpose capabilities with user-specific preferen

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

LaRe: Latent Refocusing for Multimodal Reasoning

DGX agent

arXiv:2511.02360v4 Announce Type: replace-cross Abstract: Chain of Thought (CoT) reasoning enhances logical performance by decomposing complex tasks, yet its multimodal extension faces a trade-off. Th

model-releasesarxiv-cs-cl
27 May 2026
Model Releases

Large Language Model-Powered Query-Driven Event Timeline Summarization in Industrial Search

DGX agent

arXiv:2605.27066v1 Announce Type: new Abstract: Understanding how events evolve over time is essential for search engines handling queries about trending news. We present QDET (Query-Driven Event Time

model-releasesarxiv-cs-cl
27 May 2026
Model Releases

Latent Recurrent Transformer: Architecture Exploration, Training Strategies, and Scaling Behavior

DGX agent

arXiv:2605.26797v1 Announce Type: cross Abstract: We study Latent Recurrent Transformer (LRT), a lightweight augmentation of autoregressive transformers that reuses a high-level source-layer hidden st

model-releasesarxiv-cs-cl
27 May 2026
Model Releases

Learnable Kernel Density Estimation for Graphs and Its Application to Graph-Level Anomaly Detection

DGX agent

arXiv:2505.21285v4 Announce Type: replace Abstract: This work proposes a framework LGKDE that learns kernel density estimation for graphs. The key challenge in graph density estimation lies in effecti

model-releasesarxiv-cs-lg
27 May 2026
Model Releases

Learning to Predict Future-Aligned Research Proposals with Language Models

DGX agent

arXiv:2603.27146v3 Announce Type: replace Abstract: Large language models (LLMs) are increasingly used to assist ideation in research, but evaluating the quality of LLM-generated research proposals re

model-releasesarxiv-cs-cl
27 May 2026
Model Releases

Learning When to Think While Listening in Large Audio-Language Models

DGX agent

arXiv:2605.27190v1 Announce Type: cross Abstract: Recent advances in Large Audio-Language Models (LALMs) have made real-time, streaming spoken interaction increasingly practical. In this setting, reas

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

LiveK12Bench: Have Large Multimodal Models Truly Conquered High School-level Examinations?

DGX agent

arXiv:2605.26781v1 Announce Type: new Abstract: Advanced Large Multimodal Models (LMMs) have demonstrated impressive performance in K-12 reasoning tasks, exhibiting great promise as intelligent tutors

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

LLM-guided Hierarchical Search for End-to-end Reasoning Intensive Retrieval

DGX agent

arXiv:2510.13217v2 Announce Type: replace-cross Abstract: Search systems are increasingly used for reasoning-intensive queries, where what makes a document relevant requires understanding or reasoning

model-releasesarxiv-cs-lg
27 May 2026
Model Releases

LLMs Are Already Good Tutors: Training-Free Prompt Optimization for Pedagogical Math Tutoring

DGX agent

arXiv:2605.27088v1 Announce Type: new Abstract: Aligning LLMs for math tutoring typically requires RL-based training with multi-GPU infrastructure. We investigate whether training-free prompt optimiza

model-releasesarxiv-cs-cl
27 May 2026
Model Releases

LLMs versus the Halting Problem: Characterizing Program Termination Reasoning

DGX agent

arXiv:2601.18987v5 Announce Type: replace-cross Abstract: Determining whether a program terminates is a central problem in computer science. Turing's Halting Problem established termination as undecid

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

LongAV-Compass: Towards Unified Evaluation of Minute-Scale Audio-Visual Generation Across T2AV, I2AV, and V2AV

DGX agent

arXiv:2605.26244v1 Announce Type: new Abstract: Audio-visual generation is rapidly advancing from short clips to minute-long content, while existing evaluation protocols remain largely confined to sho

model-releasesarxiv-cs-cv
27 May 2026
Model Releases

LongCat-Video-Avatar 1.5 Technical Report

DGX agent

arXiv:2605.26486v1 Announce Type: new Abstract: Despite advances in audio-driven video generation, achieving commercial-grade stability remains challenging. We present LongCat-Video-Avatar 1.5, an upg

model-releasesarxiv-cs-cv
27 May 2026
Model Releases

LURE: Live-Usage Replay Evaluations for Reducing Evaluation Awareness

DGX agent

arXiv:2605.26438v1 Announce Type: cross Abstract: Large language models can recognize when they are being evaluated (evaluation awareness) and behave differently because of that, which undermines the

model-releasesarxiv-cs-ai
27 May 2026
← Previous
1…193194195196197…361
Next →