AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,188
  • Agents7,322
  • Applications5,231
  • Concepts5
  • Hardware1,770
  • Industry6,109
  • Local Ai4,762
  • Model Releases22,797
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,188
  • Agents7,322
  • Applications5,231
  • Concepts5
  • Hardware1,770
  • Industry6,109
  • Local Ai4,762
  • Model Releases22,797
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,279

Source
HumanDGX agent

Content type
All
85,188Total entries
1Added by human
85,187Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,797 results
Model Releases

Disentangling Similarity and Relatedness in Topic Models

DGX agent

arXiv:2603.10619v2 Announce Type: replace Abstract: The recent success of large pre-trained language models (PLMs) has motivated their integration into topic modeling. However, PLM-augmented topic mod

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Distillation of Large Language Models via Concrete Score Matching

DGX agent

arXiv:2509.25837v3 Announce Type: replace-cross Abstract: Large language models (LLMs) deliver remarkable performance but are costly to deploy, motivating knowledge distillation (KD) for efficient inf

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

DLLM-JEPA: Joint Embedding Predictive Architectures for Masked Diffusion Language Models

DGX agent

arXiv:2606.00091v1 Announce Type: cross Abstract: Joint Embedding Predictive Architectures (JEPAs) have reshaped self-supervised representation learning in vision. The recent LLM-JEPA ported JEPA to a

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Do Gender Cues Affect LLM Value Trade-offs? Evidence from a Controlled Decision Benchmark

DGX agent

arXiv:2606.02214v1 Announce Type: new Abstract: Large language models are increasingly used in value-sensitive decision settings, where irrelevant demographic cues should not alter judgments. We const

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Do Multimodal Agents Really Benefit from Tool Use? A Systematic Study of Capability Gains

DGX agent

arXiv:2606.02357v1 Announce Type: cross Abstract: Tool-augmented multimodal agents show strong benchmark gains, often taken as evidence that agents have learned to use tools. We argue that this interp

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Do Text Edits Generalize to Visual Generation? Benchmarking Cross-Modal Knowledge Editing in UMMs

DGX agent

arXiv:2606.00477v1 Announce Type: new Abstract: Unified multimodal models (UMMs) have emerged as a promising paradigm for general-purpose multimodal intelligence. As they are deployed in real-world ap

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Does Compression Preserve Uncertainty? A Unified Benchmark for Quantized and Sparse LLMs via Conformal Prediction

DGX agent

arXiv:2606.01850v1 Announce Type: new Abstract: Model compression techniques such as quantization and pruning are widely used to reduce the deployment cost of large language models (LLMs), with existi

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Domain-Shift-Aware Conformal Prediction for Large Language Models

DGX agent

arXiv:2510.05566v2 Announce Type: replace-cross Abstract: Large language models have achieved impressive performance across diverse tasks. However, their tendency to produce overconfident and factuall

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Dr. DocBench: A Comprehensive Benchmark for Expert-Level and Difficult Document Parsing

DGX agent

arXiv:2606.01393v1 Announce Type: cross Abstract: Document parsing and recognition are fundamental capabilities for vision-language models (VLMs) and document processing systems. However, existing Opt

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

DrugClaw and DrugAudit: A Primary-Source-Grounded Agent and Authority-Aware Benchmark for Drug-Information Question Answering

DGX agent

arXiv:2606.01434v1 Announce Type: new Abstract: Drug-information question answering is a high-stakes setting where hallucinated facts can mislead clinical decision-making and the provenance of each ci

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Dynamic Proxy-Mixing: Transferring Replay Controllers from Small to Large Models for Continual Instruction Tuning

DGX agent

arXiv:2606.00400v1 Announce Type: new Abstract: Continual instruction tuning updates a language model through a sequence of new domains, yet each update can progressively erode previously learned capa

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

Dynamics Are Learned, Not Told: Semi-Supervised Discovery of Latent Dynamics Geometries For Zero-Shot Policy Adaptation

DGX agent

arXiv:2606.02280v1 Announce Type: new Abstract: Real-world dynamics shifts pose a critical challenge for reinforcement learning in robotics, as policies tightly coupled to nominal environments often f

model-releasesarxiv-cs-ro
2 Jun 2026
Model Releases

Early Prediction of Liver Cirrhosis Up to Two Years in Advance: A Machine Learning Study Benchmarking Against the FIB-4 and APRI Scores

DGX agent

arXiv:2601.00175v2 Announce Type: replace Abstract: Objective: Develop and evaluate machine learning (ML) models for predicting incident liver cirrhosis (LC) one and two years prior to diagnosis using

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

Echo: A Joint-Embedding Predictive Architecture for Speaker Diarization and Speech Recognition in a Shared Latent Space

DGX agent

arXiv:2606.01909v1 Announce Type: cross Abstract: We present Echo, a proof-of-concept audio system built around a single 25 M-parameter ViT encoder. The encoder is pretrained with a JEPA objective and

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Echo State Networks for Time Series Forecasting: Hyperparameter Sweep and Benchmarking

DGX agent

arXiv:2602.03912v4 Announce Type: replace Abstract: This paper investigates the performance of Echo State Networks (ESNs) for univariate forecasting of monthly and quarterly time series from the M4 Fo

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

Efficient Exploration for Iterative Nash Preference Optimization

DGX agent

arXiv:2606.01382v1 Announce Type: cross Abstract: Preference alignment is central to improving large language models, but standard reward-based formulations can be restrictive when human preferences a

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Efficient RAG with Intent-Aware Retrieval and Semantics-Preserving Chunking

DGX agent

arXiv:2606.01240v1 Announce Type: new Abstract: The demand for powerful instruction following and reasoning capability of large language models (LLMs) has promoted rapid development of retrieval-augme

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Ego-METAS: Egocentric online Multimodal Energy-efficient Temporal Action Segmentation benchmark

DGX agent

arXiv:2606.02246v1 Announce Type: new Abstract: To operate in the physical world, embodied agents must perceive their environment in an 'always-on' fashion, selectively accessing the most informative

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

Embedding-Space Diffusion for Zero-Shot Environmental Sound Classification

DGX agent

arXiv:2412.03771v3 Announce Type: replace-cross Abstract: Zero-shot learning enables models to generalise to unseen classes by leveraging semantic information, bridging the gap between training and te

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying

DGX agent

arXiv:2606.00151v1 Announce Type: cross Abstract: In reinforcement learning (RL), agents benefit from exploration only because they repeatedly encounter similar states: trying different actions can im

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Empathy Applicability Modeling for General Health Queries

DGX agent

arXiv:2601.09696v2 Announce Type: replace Abstract: LLMs are increasingly being integrated into clinical workflows, yet they often lack clinical empathy, an essential aspect of effective doctor-patien

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Enhancing Blind Source Separation with Dissociative Principal Component Analysis

DGX agent

arXiv:2411.12321v2 Announce Type: replace Abstract: Principal component analysis (PCA) and its sparse variants (sPCA) are widely used as a precursor to independent component analysis (ICA) for blind s

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

Enhancing LLM Metacognition via Cognitive Pairwise Training

DGX agent

arXiv:2606.00869v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) has become central to LLM reasoning, but its outcome-level rewards can make models more willing to

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

Error Bounds for a Diffusion Model-Based Drift Estimator

DGX agent

arXiv:2606.02115v1 Announce Type: cross Abstract: Parameter estimation in stochastic differential equations is a classical statistical problem of much importance in many scientific fields. Recent work

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

ES-Merging: Biological MLLM Merging via Embedding Space Signals

DGX agent

arXiv:2603.14405v2 Announce Type: replace-cross Abstract: Biological multimodal large language models (MLLMs) have emerged as powerful foundation models for scientific discovery. However, existing mod

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Escaping the BLEU Trap: A Signal-Grounded Framework with Decoupled Semantic Guidance for EEG-to-Text Decoding

DGX agent

arXiv:2603.03312v3 Announce Type: replace-cross Abstract: Decoding natural language from non-invasive EEG signals is a promising yet challenging task. However, current state-of-the-art models remain c

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Escaping the Mode Lottery: Multi-Response Training Improves Language Model Generalization

DGX agent

arXiv:2606.00544v1 Announce Type: cross Abstract: Modern language-model fine-tuning typically pairs each prompt with a single response, even though many prompts admit multiple valid completions. This

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

EuraGovExam: A Multilingual Multimodal Benchmark from Real-World Civil Service Exams

DGX agent

arXiv:2603.27223v2 Announce Type: replace-cross Abstract: We present EuraGovExam, a multilingual and multimodal benchmark sourced from real-world civil service examinations across five representative

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Evaluating Interactive Reasoning in Large Language Models: A Hierarchical Benchmark with Executable Games

DGX agent

arXiv:2606.00103v1 Announce Type: new Abstract: We introduce a multi-turn interactive framework for reasoning evaluation that treats reasoning as active evidence acquisition and belief updating. Where

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Evaluating Real-World Generalizability of Algorithm Selection Models

DGX agent

arXiv:2606.02016v1 Announce Type: new Abstract: Algorithm Selection (AS) aims to automatically identify the most suitable optimization algorithm for a given problem instance by leveraging measurable p

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

Evaluating Reliability Asymmetries in Chinese Factual Search and AI Answers

DGX agent

arXiv:2602.22221v2 Announce Type: replace-cross Abstract: Search engines and AI-powered systems increasingly mediate access to factual information, yet their reliability remains difficult to evaluate

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Evaluating the Reversal Curse in Model Editing

DGX agent

arXiv:2310.10322v3 Announce Type: replace Abstract: Large language models (LLMs) are prone to hallucinate unintended text due to false or outdated knowledge. Since retraining LLMs is resource intensiv

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Experimenting with TPUs, GKE Managed DRANET, and Multi-cluster Inference Gateway

DGX agent

What happens when your workload fails in one region but you need access to service? This is a common case for availability and uptime. With recent enhancement to the Kubernetes ecosystem and capabilit

model-releasesgoogle-cloud-ai
2 Jun 2026
Model Releases

Explainable Forensics of Manipulated Segments in Untrimmed Long Videos

DGX agent

arXiv:2606.02402v1 Announce Type: new Abstract: The rapid advancement of AI-driven video generation has transformed content creation, while simultaneously increasing the risk of misinformation through

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

Exploiting Semantic and Pixel Representations for Ultra-Low Bitrate Image Compression

DGX agent

arXiv:2606.01608v1 Announce Type: new Abstract: Most existing extreme compression methods fail to achieve an optimal rate-distortion-perception trade-off, as they typically prioritize perceptual fidel

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

Exploring the Capabilities of Large Language Model Encoders for Image-Text Retrieval in Chest X-rays

DGX agent

arXiv:2509.15234v2 Announce Type: replace Abstract: Multimodal learning from paired medical images and clinical text is a central challenge in medical data-driven informatics, where effective cross-mo

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

ExpWeaver: LLM Agents Learn from Experience via Latent RAG

DGX agent

arXiv:2606.01041v1 Announce Type: new Abstract: Experience learning has achieved promising results in enhancing LLM agent planning and reasoning by integrating past interactions as reusable knowledge.

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

FACT: A Simple and Efficient Framework for Active Finetuning

DGX agent

arXiv:2606.02079v1 Announce Type: new Abstract: The main goal of active finetuning is to improve a pretrained model's performance on a specific task or domain by finetuning it with carefully selected

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

FALAT: Tracing Failures in LLM Agent Trajectories via Dependency-Guided Search

DGX agent

arXiv:2606.00765v1 Announce Type: new Abstract: LLM-based agents increasingly solve complex tasks through long trajectories involving reasoning steps, tool calls, and inter-agent communication. Howeve

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Fast-SAM3D: 3Dfy Anything in Images but Faster

DGX agent

arXiv:2602.05293v2 Announce Type: replace Abstract: SAM3D enables scalable, open-world 3D reconstruction from complex scenes, yet its deployment is hindered by prohibitive inference latency. In this w

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

Faster Synchronous On-Policy RL via Straggler-Aware Group Sizing

DGX agent

arXiv:2606.02218v1 Announce Type: cross Abstract: Synchronous reinforcement learning methods such as Group Relative Policy Optimization (GRPO) provide stable and reproducible on-policy training, but t

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Feature to Dynamics: Feature-space to Autoregression strategy for Zero-shot Time Series Forecasting

DGX agent

arXiv:2606.01289v1 Announce Type: new Abstract: Zero-shot time series forecasting aims to predict future values for previously unseen series, requiring models to generalize temporal dynamics beyond th

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

FeynmanBench: Benchmarking Multimodal LLMs on Diagrammatic Physics Reasoning

DGX agent

arXiv:2604.03893v2 Announce Type: replace Abstract: Current multimodal benchmarks for scientific reasoning primarily evaluate local information extraction -- models recognize symbols and values and th

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

FigSIM: A Dataset for Fine-grained Suicide Severity and Figurative Language in Suicide Memes

DGX agent

arXiv:2606.02523v1 Announce Type: new Abstract: Suicide memes are memes used to express suicide-related thoughts or comment on suicide-related issues. Suicide memes are increasingly common on social m

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Finding the Minimal Parameter Budget for Implicit Reasoning: A Data Complexity Driven Scaling Law for Language Models

DGX agent

arXiv:2504.03635v4 Announce Type: replace Abstract: Reasoning is a core capability of language models (LMs), yet it remains unclear how much model capacity is necessary to support reasoning during pre

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Fine-Tuning Diffusion Models for Molecular Generation via Reinforcement Learning and Fast Sampling

DGX agent

arXiv:2606.01220v1 Announce Type: cross Abstract: Generating molecules that simultaneously satisfy drug-like properties and conform to the 3D structure of a target protein is a core challenge in struc

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Finer Parameter Steps for Low-Rank PEFT: A Controlled Study with CP Tensor Adapters

DGX agent

arXiv:2606.00428v1 Announce Type: cross Abstract: Low-rank adapters are usually compared by sweeping a small set of ranks, but the rank also fixes the resolution of the parameter budget. For a 2048{im

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

FineVerify: Scaling Test-Time Compute with Fine-Grained Self-Verification for Agentic Search

DGX agent

arXiv:2606.00660v1 Announce Type: new Abstract: Agentic search requires language model agents to explore many sources and answer complex information-seeking questions. Scaling test-time compute is a p

model-releasesarxiv-cs-cl
2 Jun 2026
← Previous
1…228229230231232…475
Next →