AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries90,914
  • Agents7,747
  • Applications5,533
  • Concepts5
  • Hardware1,920
  • Industry6,196
  • Local Ai5,094
  • Model Releases24,727
  • Research20,781
  • Safety13,739
  • Syntheses17
  • Tools1,678
  • Tutorials3,477

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries90,914
  • Agents7,747
  • Applications5,533
  • Concepts5
  • Hardware1,920
  • Industry6,196
  • Local Ai5,094
  • Model Releases24,727
  • Research20,781
  • Safety13,739
  • Syntheses17
  • Tools1,678
  • Tutorials3,477

Source
HumanDGX agent

Content type
90,914Total entries
1Added by human
90,913Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
53,690 results
Model Releases

Decoupled Training with Local Reinforcement Fine-Tuning in Federated Learning

DGX agent

arXiv:2605.27900v1 Announce Type: new Abstract: Federated Learning (FL) with pre-trained Vision-Language Models (VLMs) has emerged as a promising paradigm for various downstream tasks. By leveraging i

model-releasesarxiv-cs-cv
28 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Deformable Gaussian Occupancy: Decoupling Rigid and Nonrigid Motion with Factorized Distillation

DGX agent

arXiv:2605.28587v1 Announce Type: new Abstract: Understanding dynamic 3D environments is essential for safe autonomous driving, particularly when reasoning about human-centric, nonrigid agents. Howeve

model-releasesarxiv-cs-cv
28 May 2026
Model Releases

Differential syntactic and semantic encoding in LLMs

DGX agent

arXiv:2601.04765v4 Announce Type: replace-cross Abstract: We study how syntactic and semantic information is encoded in inner layer representations of Large Language Models (LLMs), focusing on the ver

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Efficient and Scalable Provenance Tracking for LLM-Generated Code Snippets

DGX agent

arXiv:2605.28510v1 Announce Type: cross Abstract: Large language models (LLMs) for code completion and generation are increasingly used in software development, yet they may reproduce training example

model-releasesarxiv-cs-ai
28 May 2026
Research

Fine-Tuning Dynamics of In-Context Factual Recall in Transformers

DGX agent

arXiv:2605.27774v1 Announce Type: new Abstract: In-context learning -- performing tasks based on examples given in the prompt -- is an important capability that has emerged in large language models an

researcharxiv-cs-lg
28 May 2026
Model Releases

From Knowing to Doing: A Memory-Controlled Benchmark for LLM Trading Agents on Stock Markets

DGX agent

arXiv:2605.28359v1 Announce Type: new Abstract: Evaluating whether large language model (LLM) agents can profit in capital markets is increasingly framed as end-to-end trading: place an agent in a his

model-releasesarxiv-cs-ai
28 May 2026
Applications

GEM: Generative Supervision Helps Embodied Intelligence

DGX agent

arXiv:2605.28548v1 Announce Type: new Abstract: Embodied Vision-Language Models (VLMs) have demonstrated impressive performance and generalization in robotics, particularly within Vision-Language-Acti

applicationsarxiv-cs-cv
28 May 2026
Model Releases

Hurwitz Quaternion Multiplicative Quantization for KV Cache Compression

DGX agent

arXiv:2605.27646v1 Announce Type: cross Abstract: We propose extbf{Hurwitz Quaternion Multiplicative Quantization (HQMQ)}, a extbf{calibration-free} method for KV cache compression of large language m

model-releasesarxiv-cs-ai
28 May 2026
Tutorials

Identifiable Bayesian Deep Generative Copulas with Unknown Layer Widths for Data with Arbitrary Marginal Distributions

DGX agent

arXiv:2605.27523v1 Announce Type: cross Abstract: Deep generative models offer powerful tools for multivariate data analysis, but their black-box architectures are often unidentified and difficult to

tutorialsarxiv-cs-lg
28 May 2026
Model Releases

IRDS: Interpretable RLVR Data Selection via Verifier-Coupled Sparse Autoencoder Coverage

DGX agent

arXiv:2605.28247v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has become a key technique for en- hancing LLM reasoning, yet its data ineffi- ciency remains a

model-releasesarxiv-cs-ai
28 May 2026
Research

La-Proteina: Atomistic Protein Generation via Partially Latent Flow Matching

DGX agent

arXiv:2507.09466v2 Announce Type: replace Abstract: Recently, many generative models for de novo protein structure design have emerged. Yet, only few tackle the difficult task of directly generating f

researcharxiv-cs-lg
28 May 2026
Research

Latent Diffusion for Missing Data

DGX agent

arXiv:2605.28427v1 Announce Type: new Abstract: Diffusion models have emerged as powerful generative approaches for missing-data imputation, yet most existing methods operate directly in data space an

researcharxiv-cs-lg
28 May 2026
Model Releases

Learning to Translate from Soft to Hard LLM Prompts

DGX agent

arXiv:2605.27642v1 Announce Type: new Abstract: Soft prompt tuning is a parameter-efficient method for adapting LLMs to specific tasks, but suffers from a lack of interpretability. Building on recent

model-releasesarxiv-cs-cl
28 May 2026
Model Releases

LLM Zeroth-Order Fine-Tuning is an Inference Workload

DGX agent

arXiv:2605.28760v1 Announce Type: new Abstract: Zeroth-order (ZO) fine-tuning is attractive for large language models because it replaces backpropagation with forward objective evaluations. Existing i

model-releasesarxiv-cs-lg
28 May 2026
Model Releases

MIRAGE: Context-Aware Prompt Injection against Mobile GUI Agents via User-Generated Content

DGX agent

arXiv:2605.28116v1 Announce Type: cross Abstract: Mobile graphical user interface (GUI) agents driven by vision-language models (VLMs) perceive the screen as rendered pixels and choose actions from wh

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

NanoVDR: Distilling a 2B Vision-Language Retriever into a 70M Text-Only Encoder for Visual Document Retrieval

DGX agent

arXiv:2603.12824v2 Announce Type: replace-cross Abstract: Vision-Language Model (VLM) based retrievers have advanced visual document retrieval (VDR) to impressive quality. They require the same multi-

model-releasesarxiv-cs-cv
28 May 2026
Model Releases

OccuReward: LLM-Guided Occupant-Centric Reward Shaping for Demographic Equity in Grid-Interactive Buildings

DGX agent

arXiv:2605.28168v1 Announce Type: new Abstract: Large language models (LLMs) have demonstrated promising capability in generating reward functions for deep reinforcement learning (DRL)-based building

model-releasesarxiv-cs-ai
28 May 2026
Safety

OGER: A Robust Offline-Guided Exploration Reward for Hybrid Reinforcement Learning

DGX agent

arXiv:2604.18530v2 Announce Type: replace Abstract: Recent advancements in Reinforcement Learning with Verifiable Rewards (RLVR) have significantly improved Large Language Model (LLM) reasoning, yet m

safetyarxiv-cs-ai
28 May 2026
Local Ai

OmniVerifier-M1: Multimodal Meta-Verifier with Explicit Structured Recalibration

DGX agent

arXiv:2605.28805v1 Announce Type: cross Abstract: Visual outcomes are increasingly central to multimodal large language models, making reliable and fine-grained verification essential for scaling gene

local-aiarxiv-cs-ai
28 May 2026
Model Releases

On Compositional Learning Behaviours in Formal Mathematics

DGX agent

arXiv:2605.28512v1 Announce Type: new Abstract: Self-evolving scientific agents capable of conquering the hard tail of formal mathematics require Compositional Learning Behaviours (CLBs) -- the capaci

model-releasesarxiv-cs-cl
28 May 2026
Model Releases

OralAgent: Integrating Reasoning, Tools, and Knowledge for Interactive Dental Image Analysis

DGX agent

arXiv:2605.27378v1 Announce Type: new Abstract: Dental image analysis plays a pivotal role in supporting accurate diagnosis and treatment planning in oral healthcare. Although recent advances have pro

model-releasesarxiv-cs-cl
28 May 2026
Research

PEAR: Equal Area Weather Forecasting on the Sphere

DGX agent

arXiv:2505.17720v3 Announce Type: replace Abstract: Artificial intelligence is rapidly reshaping the natural sciences, with weather forecasting emerging as a flagship AI4Science application where mach

researcharxiv-cs-lg
28 May 2026
Model Releases

PEFT-Arena: Understanding Parameter-Efficient Finetuning from a Stability-Plasticity Perspective

DGX agent

arXiv:2605.28819v1 Announce Type: cross Abstract: Parameter-efficient finetuning (PEFT) has become the standard approach for adapting large language models, yet evaluations largely emphasize downstrea

model-releasesarxiv-cs-cl
28 May 2026
Model Releases

PINE: Pruning Boosted Tree Ensembles with Conformal In-Distribution Prediction Equivalence

DGX agent

arXiv:2605.28068v1 Announce Type: new Abstract: Tree ensembles are machine learning models with strong predictive performance and interpretability, and remain widely used for tabular data. Standard pr

model-releasesarxiv-cs-lg
28 May 2026
Model Releases

PointQ-Bench: Benchmarking Diagnostic and Interpretable Point Cloud Quality Assessment

DGX agent

arXiv:2605.28241v1 Announce Type: new Abstract: Point cloud quality plays a critical role in 3D acquisition, reconstruction, rendering, and perception, yet existing point cloud quality assessment (PCQ

model-releasesarxiv-cs-cv
28 May 2026
Model Releases

PromptEmbedder:: Efficient and Transferable Text Embedding via Dual-LLM Soft Prompting

DGX agent

arXiv:2605.28066v1 Announce Type: cross Abstract: Large Language Models (LLMs) have demonstrated remarkable efficacy in text embedding, yet current adaptation methods like LoRA face significant bottle

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Qwen-Image-Bench: From Generation to Creation in Text-to-Image Evaluation

DGX agent

arXiv:2605.28091v1 Announce Type: new Abstract: Text-to-Image generation has evolved from basic image synthesis into a frequently used core capability in professional creative workflows, where simple

model-releasesarxiv-cs-cv
28 May 2026
Model Releases

Reflective Dialogue between Teacher and Solver Agents for Video Question Answering

DGX agent

arXiv:2605.27885v1 Announce Type: new Abstract: Various approaches have been proposed to adapt Vision-Language Models (VLMs) to specialized domains for Video Question Answering, including fine-tuning

model-releasesarxiv-cs-cv
28 May 2026
Agents

ResearchMath-14K: Scaling Research-Level Mathematics via Agents

DGX agent

arXiv:2605.28003v1 Announce Type: new Abstract: The frontier of mathematics is defined by problems whose solutions are not yet known, yet it remains unclear whether language models can meaningfully en

agentsarxiv-cs-cl
28 May 2026
Model Releases

Resolution-free neural surrogates for geometric parameterization and mapping with spatially varying fields

DGX agent

arXiv:2605.28551v1 Announce Type: new Abstract: Many imaging problems require computing spatial transformations induced by spatially varying intensity, feature, or density fields. Canonical examples i

model-releasesarxiv-cs-cv
28 May 2026
Model Releases

SAME: Stabilized Mixture-of-Experts for Multimodal Continual Instruction Tuning

DGX agent

arXiv:2602.01990v2 Announce Type: replace-cross Abstract: Multimodal Large Language Models (MLLMs) achieve strong performance through instruction tuning, but real-world deployment requires them to con

model-releasesarxiv-cs-ai
28 May 2026
Safety

SEMAGIC: Learning Semantically Consistent Deformable 3D Representations from In-the-Wild Images

DGX agent

arXiv:2605.27938v1 Announce Type: new Abstract: Learning deformable 3D object models from single-view in-the-wild images has enabled impressive 3D shape reconstruction without supervision. However, it

safetyarxiv-cs-cv
28 May 2026
Safety

Singular Vectors of Attention Heads Align with Features

DGX agent

arXiv:2602.13524v2 Announce Type: replace-cross Abstract: Identifying feature representations in language models is a central task in mechanistic interpretability. Several recent studies have made the

safetyarxiv-cs-ai
28 May 2026
Model Releases

StoryMI: Steerable Multi-Agent Therapeutic Dialogue Generation

DGX agent

arXiv:2605.27393v1 Announce Type: cross Abstract: Large language models (LLMs) can generate fluent dialogue, but prior works lack situational grounding, dynamic strategy control, and evaluation aligne

model-releasesarxiv-cs-ai
28 May 2026
Research

The Future of Facts: Tracing the Factual Generation-Verification Gap

DGX agent

arXiv:2605.27564v1 Announce Type: cross Abstract: Language models are becoming the default interface to factual knowledge, yet they often verify outputs more reliably than they generate them. This gen

researcharxiv-cs-ai
28 May 2026
Safety

The Obfuscation Atlas: Mapping Where Honesty Emerges in RLVR with Deception Probes

DGX agent

arXiv:2602.15515v2 Announce Type: replace-cross Abstract: Training against white-box deception detectors has been proposed as a way to make AI systems honest. However, such training risks models learn

safetyarxiv-cs-ai
28 May 2026
Model Releases

TRACER: Turn-level Regret Matching with Inner Reinforcement Credit for Cooperative Multi-LLM Reasoning

DGX agent

arXiv:2605.28699v1 Announce Type: new Abstract: Large language models increasingly rely on either reinforcement learning or multi-agent prompting to improve reasoning, yet these two paradigms remain d

model-releasesarxiv-cs-ai
28 May 2026
Research

Why We Need Speech to Evaluate Speech Translation

DGX agent

arXiv:2605.28227v1 Announce Type: new Abstract: Speech translation models are increasingly capable of preserving speech-specific information (e.g., speaker gender, prosody, and emphasis), yet evaluati

researcharxiv-cs-cl
28 May 2026
Model Releases

A Dataset of Robot-Patient and Doctor-Patient Medical Dialogues for Spoken Language Processing Tasks

DGX agent

arXiv:2605.26747v1 Announce Type: new Abstract: Large Language Models (LLMs) have brought huge improvements to Artificial Intelligence (AI), which can be applied to general-purpose tasks. However, the

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

Attribute-Based Diagnosis of LLM Alignment with Hate Speech Annotations

DGX agent

arXiv:2605.27025v1 Announce Type: new Abstract: Hate speech annotation is costly, subjective, and prone to annotator disagreement, making large-scale dataset construction challenging. We systematicall

model-releasesarxiv-cs-cl
27 May 2026
Model Releases

Chain Of Thought Compression: A Theoretical Analysis

DGX agent

arXiv:2601.21576v2 Announce Type: replace Abstract: Chain-of-Thought (CoT) has unlocked advanced reasoning abilities of Large Language Models (LLMs) with intermediate steps, yet incurs prohibitive com

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

CktGen: Automated Analog Circuit Design with Generative Artificial Intelligence

DGX agent

arXiv:2410.00995v3 Announce Type: replace Abstract: The automatic synthesis of analog circuits presents significant challenges. Most existing approaches formulate the problem as a single-objective opt

model-releasesarxiv-cs-lg
27 May 2026
Model Releases

Composition Collapse: Stable Factual Knowledge Does Not Imply Compositional Reasoning

DGX agent

arXiv:2605.26789v1 Announce Type: new Abstract: Post-training is routinely evaluated through aggregate benchmark scores that treat multi-hop reasoning as a single capability -- as if a model that answ

model-releasesarxiv-cs-ai
27 May 2026
Safety

CroCo: Cross-Lingual Contrastive Preference Tuning on Self-Generations

DGX agent

arXiv:2605.26293v1 Announce Type: cross Abstract: Prior work establishes that controlled contrastiveness between self-generated responses from large language models, set via reward scores, improves do

safetyarxiv-cs-ai
27 May 2026
Model Releases

Developing a Totally Unimodular Linear Program for Optimal Conformance Checking: When and Why It Complements A*

DGX agent

arXiv:2605.26938v1 Announce Type: new Abstract: Alignment-based conformance checking is the state-of-the-art approach for comparing observed process executions with normative process models. The stand

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

Faithfulness Evaluation for Decoder-only LLM Attributions with Controlled Retained Information

DGX agent

arXiv:2601.03089v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are increasingly evaluated with input attribution methods, yet comparing such explanations remains challenging. E

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

FineVLA: Fine-Grained Instruction Alignment for Steerable Vision-Language-Action Policies

DGX agent

arXiv:2605.27284v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models are increasingly expected to not only complete robot tasks, but also follow human instructions about how those tas

model-releasesarxiv-cs-ai
27 May 2026
Safety

Furina: Fragmented Uncertainty-Driven Refusal Instability Attack

DGX agent

arXiv:2605.26158v1 Announce Type: cross Abstract: Safety alignment in large language models (LLMs) and multimodal large language models (MLLMs) is commonly assumed to operate as a near-binary threshol

safetyarxiv-cs-ai
27 May 2026
← Previous
1…438439440441442…1119
Next →