AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,630Total entries
1Added by human
84,629Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,288 results
Model Releases

ContinuousBench: Can Differentially Private Synthetic Text Improve Capabilities?

DGX agent

arXiv:2606.01849v1 Announce Type: cross Abstract: Differentially private (DP) text synthesis promises to unlock sensitive corpora for model training, but it remains unclear whether DP synthetic data t

model-releasesarxiv-cs-cl
2 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Controllable Value Alignment in Large Language Models through Neuron-Level Editing

DGX agent

arXiv:2602.07356v2 Announce Type: replace Abstract: Aligning large language models (LLMs) with human values has become increasingly important as their influence on human behavior and decision-making e

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

Correcting Gradient-Based Circuit Localization via Interaction-Aware Backpropagation

DGX agent

arXiv:2505.17630v4 Announce Type: replace Abstract: Circuit localization methods aim to identify the subset of model components responsible for specific behaviors in large language models, enabling de

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

CRAB-Bench: Evaluating LLM Agents under Complex Task Dependencies and Human-aligned User Simulation

DGX agent

arXiv:2606.01815v1 Announce Type: new Abstract: Evaluating LLM agents in realistic service scenarios requires complex task dependencies, imperfect user behavior, and an evaluation that accommodates mu

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

CRAM: Centroid-Routing and Adaptive MoE for Multimodal Continual Instruction Tuning

DGX agent

arXiv:2606.02502v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) unify heterogeneous vision-language tasks under a shared generative framework via instruction tuning, yet real-

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

CRMA: A Spectrally-Bounded Backbone for Modular Continual Fine-Tuning of LLMs

DGX agent

arXiv:2606.00382v1 Announce Type: new Abstract: Sequential fine-tuning of large language models forces a choice: let the shared substrate keep learning and accept catastrophic forgetting, or freeze it

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

Cross-Environment Neural Reranking for Sample-Efficient Action Selection in Text-Based Agents

DGX agent

arXiv:2606.02204v1 Announce Type: new Abstract: Large language model agents achieve strong performance on text-based benchmarks but incur prohibitive inference costs, motivating the use of compact neu

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Cross-Generational Transfer of Adversarial Attacks Reveals Non-Monotonic Safety Alignment in LLMs

DGX agent

arXiv:2606.00813v1 Announce Type: cross Abstract: Safety alignment in LLMs does not improve monotonically across model generations. Studying four generations of Google's Gemma family (7B-31B) with qua

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

CSRP: Chain-of-Thought Reasoning for Chinese Text Correction via Reinforcement Learning with Efficiency-Aware Rewards

DGX agent

arXiv:2606.00020v1 Announce Type: cross Abstract: Large Language Model (LLM) based Chinese Grammatical Error Correction (CGEC) systems face two critical challenges: general-purpose models lack special

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

CultureForest: Understanding and Evaluating Cultural Norm Grounded Reasoning in LLMs

DGX agent

arXiv:2606.01879v1 Announce Type: new Abstract: Existing research largely reduces cultural intelligence in LLMs to a knowledge-level problem, overlooking whether models can effectively utilize their a

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

CV-Arena: An Open Benchmark for Instructional Computer Vision Problem Solving with Human-AI Collaborative Preferences

DGX agent

arXiv:2606.00931v1 Announce Type: cross Abstract: Instruction-guided image editing is becoming a general interface for visual work, yet existing benchmarks still focus largely on narrow appearance edi

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

DAG-MoE: From Simple Mixture to Structural Aggregation in Mixture-of-Experts

DGX agent

arXiv:2606.01062v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) models have become a leading approach for decoupling parameter count from computational cost in large language models, yet effe

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

DAG-Plan: Generating Directed Acyclic Dependency Graphs for Dual-Arm Cooperative Planning

DGX agent

arXiv:2406.09953v4 Announce Type: replace-cross Abstract: Dual-arm robots promise greater efficiency but require planning for complex tasks with nonlinear sub-task dependencies. Current methods using

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

DASH: Dual-Branch Score Distillation for Guidance-Calibrated Compact Diffusion Models

DGX agent

arXiv:2606.00798v1 Announce Type: cross Abstract: Parameter compression of class-conditional diffusion models reveals an underexplored limitation in output-level distillation: the unconditional score

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

DAStatFormer: A Hybrid Multibranch Transformer with Statistical Feature Integration for DAS-Based Pattern Recognitions

DGX agent

arXiv:2606.00081v1 Announce Type: cross Abstract: Distributed Acoustic Sensing (DAS) enables large-scale monitoring through optical fibers, but its high dimensionality and complex spatio-temporal patt

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Data Collection for Training Quality-Control AI in Carpet Manufacturing

DGX agent

arXiv:2606.01023v1 Announce Type: cross Abstract: Visual inspection remains the dominant quality-control practice in woven and tufted carpet production, yet it is slow, subjective, and inconsistent at

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Decentralized Instruction Tuning: Conflict-Aware Splitting and Weight Merging

DGX agent

arXiv:2606.01717v1 Announce Type: new Abstract: Instruction tuning aligns large language models, including multimodal ones, with diverse user intents, but scaling to heterogeneous mixtures is hindered

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

Decision-Focused On-Policy Learning for Contextual Linear Optimization with Partial Feedback

DGX agent

arXiv:2606.01081v1 Announce Type: new Abstract: Decision-focused learning (DFL) trains predictive models by optimizing downstream decision quality rather than standalone prediction accuracy. For conte

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

DECK: A Consistency x Confidence Taxonomy of LLM Hallucinations

DGX agent

arXiv:2606.02289v1 Announce Type: new Abstract: Existing hallucination taxonomies classify LLM errors by what is wrong with the output -- memorised misconceptions, reasoning failures, fluent fabricati

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Deep Research as Rubric for Reinforcement Learning

DGX agent

arXiv:2606.01091v1 Announce Type: new Abstract: Open-ended reasoning and long-form generation tasks lack reliable automatic verification signals for reward-based policy optimization. Rubrics offer a p

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Deformable Wiener Filter for Future Video Coding

DGX agent

arXiv:2606.01576v1 Announce Type: new Abstract: In-loop filters have attracted increasing attention due to the remarkable noise-reduction capability in the hybrid video coding framework. However, the

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

Density-Aware Translation of Spurious Correlations in Zero-Shot VLMs

DGX agent

arXiv:2606.01710v1 Announce Type: new Abstract: Vision-Language models (VLMs), such as CLIP, achieve powerful zero-shot classification. However, their predictions remain sensitive to spurious correlat

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

Design-MLLM: A Reinforcement Alignment Framework for Verifiable and Aesthetic Interior Design

DGX agent

arXiv:2603.13312v2 Announce Type: replace-cross Abstract: Interior design is a requirements-to-visual-plan generation process that must simultaneously satisfy verifiable spatial feasibility and compar

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

DetailMaster: Can Your Text-to-Image Model Handle Long Prompts?

DGX agent

arXiv:2505.16915v3 Announce Type: replace-cross Abstract: While recent Text-to-Image (T2I) models show impressive capabilities in synthesizing images from brief descriptions, they struggle with the lo

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Diagnosing LLM Arbitration Behavior over Pre-evidence Epistemic States in RAG-based Fact-Checking

DGX agent

arXiv:2606.01120v1 Announce Type: new Abstract: In RAG-based fact-checking, LLMs are increasingly used as verifiers to check given claims against retrieved evidence. Their parametric knowledge can ind

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Differentially Private Datastore Generation for Retrieval-Augmented Inference

DGX agent

arXiv:2606.01413v1 Announce Type: cross Abstract: It is crucial for modern on-device AI systems that rely on retrieval-augmented inference to release and share datastores without compromising individu

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

DINO-GFSA: Geo-Localization via Semantic Gated Fusion and Mamba-based Sequential Aggregation

DGX agent

arXiv:2606.00784v1 Announce Type: new Abstract: Cross-view geo-localization (CVGL) is critical for Unmanned Aerial Vehicle (UAV) self-positioning and target localization in GNSS-denied environments. H

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

Disentanglement-Based Equivariant Learning for Compositional VQA

DGX agent

arXiv:2606.02168v1 Announce Type: new Abstract: Compositional visual question answering (VQA) represents a challenging yet fundamental task that requires models to comprehend novel combinations of pre

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

Disentangling Similarity and Relatedness in Topic Models

DGX agent

arXiv:2603.10619v2 Announce Type: replace Abstract: The recent success of large pre-trained language models (PLMs) has motivated their integration into topic modeling. However, PLM-augmented topic mod

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Distillation of Large Language Models via Concrete Score Matching

DGX agent

arXiv:2509.25837v3 Announce Type: replace-cross Abstract: Large language models (LLMs) deliver remarkable performance but are costly to deploy, motivating knowledge distillation (KD) for efficient inf

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

DLLM-JEPA: Joint Embedding Predictive Architectures for Masked Diffusion Language Models

DGX agent

arXiv:2606.00091v1 Announce Type: cross Abstract: Joint Embedding Predictive Architectures (JEPAs) have reshaped self-supervised representation learning in vision. The recent LLM-JEPA ported JEPA to a

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Do Gender Cues Affect LLM Value Trade-offs? Evidence from a Controlled Decision Benchmark

DGX agent

arXiv:2606.02214v1 Announce Type: new Abstract: Large language models are increasingly used in value-sensitive decision settings, where irrelevant demographic cues should not alter judgments. We const

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Do Multimodal Agents Really Benefit from Tool Use? A Systematic Study of Capability Gains

DGX agent

arXiv:2606.02357v1 Announce Type: cross Abstract: Tool-augmented multimodal agents show strong benchmark gains, often taken as evidence that agents have learned to use tools. We argue that this interp

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Do Text Edits Generalize to Visual Generation? Benchmarking Cross-Modal Knowledge Editing in UMMs

DGX agent

arXiv:2606.00477v1 Announce Type: new Abstract: Unified multimodal models (UMMs) have emerged as a promising paradigm for general-purpose multimodal intelligence. As they are deployed in real-world ap

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Does Compression Preserve Uncertainty? A Unified Benchmark for Quantized and Sparse LLMs via Conformal Prediction

DGX agent

arXiv:2606.01850v1 Announce Type: new Abstract: Model compression techniques such as quantization and pruning are widely used to reduce the deployment cost of large language models (LLMs), with existi

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Domain-Shift-Aware Conformal Prediction for Large Language Models

DGX agent

arXiv:2510.05566v2 Announce Type: replace-cross Abstract: Large language models have achieved impressive performance across diverse tasks. However, their tendency to produce overconfident and factuall

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Dr. DocBench: A Comprehensive Benchmark for Expert-Level and Difficult Document Parsing

DGX agent

arXiv:2606.01393v1 Announce Type: cross Abstract: Document parsing and recognition are fundamental capabilities for vision-language models (VLMs) and document processing systems. However, existing Opt

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

DrugClaw and DrugAudit: A Primary-Source-Grounded Agent and Authority-Aware Benchmark for Drug-Information Question Answering

DGX agent

arXiv:2606.01434v1 Announce Type: new Abstract: Drug-information question answering is a high-stakes setting where hallucinated facts can mislead clinical decision-making and the provenance of each ci

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Dynamic Proxy-Mixing: Transferring Replay Controllers from Small to Large Models for Continual Instruction Tuning

DGX agent

arXiv:2606.00400v1 Announce Type: new Abstract: Continual instruction tuning updates a language model through a sequence of new domains, yet each update can progressively erode previously learned capa

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

Dynamics Are Learned, Not Told: Semi-Supervised Discovery of Latent Dynamics Geometries For Zero-Shot Policy Adaptation

DGX agent

arXiv:2606.02280v1 Announce Type: new Abstract: Real-world dynamics shifts pose a critical challenge for reinforcement learning in robotics, as policies tightly coupled to nominal environments often f

model-releasesarxiv-cs-ro
2 Jun 2026
Model Releases

Early Prediction of Liver Cirrhosis Up to Two Years in Advance: A Machine Learning Study Benchmarking Against the FIB-4 and APRI Scores

DGX agent

arXiv:2601.00175v2 Announce Type: replace Abstract: Objective: Develop and evaluate machine learning (ML) models for predicting incident liver cirrhosis (LC) one and two years prior to diagnosis using

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

Echo: A Joint-Embedding Predictive Architecture for Speaker Diarization and Speech Recognition in a Shared Latent Space

DGX agent

arXiv:2606.01909v1 Announce Type: cross Abstract: We present Echo, a proof-of-concept audio system built around a single 25 M-parameter ViT encoder. The encoder is pretrained with a JEPA objective and

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Echo State Networks for Time Series Forecasting: Hyperparameter Sweep and Benchmarking

DGX agent

arXiv:2602.03912v4 Announce Type: replace Abstract: This paper investigates the performance of Echo State Networks (ESNs) for univariate forecasting of monthly and quarterly time series from the M4 Fo

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

Efficient Exploration for Iterative Nash Preference Optimization

DGX agent

arXiv:2606.01382v1 Announce Type: cross Abstract: Preference alignment is central to improving large language models, but standard reward-based formulations can be restrictive when human preferences a

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Efficient RAG with Intent-Aware Retrieval and Semantics-Preserving Chunking

DGX agent

arXiv:2606.01240v1 Announce Type: new Abstract: The demand for powerful instruction following and reasoning capability of large language models (LLMs) has promoted rapid development of retrieval-augme

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Ego-METAS: Egocentric online Multimodal Energy-efficient Temporal Action Segmentation benchmark

DGX agent

arXiv:2606.02246v1 Announce Type: new Abstract: To operate in the physical world, embodied agents must perceive their environment in an 'always-on' fashion, selectively accessing the most informative

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

Embedding-Space Diffusion for Zero-Shot Environmental Sound Classification

DGX agent

arXiv:2412.03771v3 Announce Type: replace-cross Abstract: Zero-shot learning enables models to generalise to unseen classes by leveraging semantic information, bridging the gap between training and te

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying

DGX agent

arXiv:2606.00151v1 Announce Type: cross Abstract: In reinforcement learning (RL), agents benefit from exploration only because they repeatedly encounter similar states: trying different actions can im

model-releasesarxiv-cs-ai
2 Jun 2026
← Previous
1…165166167168169…361
Next →