AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,288 results
Model Releases

Adaptive Regularization for Sparsity Control in Bregman-Based Optimizers

DGX agent

arXiv:2605.07892v1 Announce Type: new Abstract: Sparse training reduces the memory and computational costs of deep neural networks. However, sparse optimization methods, e.g., those adding an ell_1 pe

model-releasesarxiv-cs-lg
11 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

AgentEscapeBench: Evaluating Out-of-Domain Tool-Grounded Reasoning in LLM Agents

DGX agent

arXiv:2605.07926v1 Announce Type: new Abstract: As LLM-based agents increasingly rely on external tools, it is important to evaluate their ability to sustain tool-grounded reasoning beyond familiar wo

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Agentick: A Unified Benchmark for General Sequential Decision-Making Agents

DGX agent

arXiv:2605.06869v1 Announce Type: new Abstract: AI agent research spans a wide spectrum: from RL agents that learn from scratch to foundation model agents that leverage pre-trained knowledge, yet no u

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

AI CFD Scientist: Toward Open-Ended Computational Fluid Dynamics Discovery with Physics-Aware AI Agents

DGX agent

arXiv:2605.06607v2 Announce Type: replace-cross Abstract: Recent LLM-based agents have closed substantial portions of the scientific discovery loop in software-only machine-learning research, in chemi

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Amortized Molecular Optimization via Group Relative Policy Optimization

DGX agent

arXiv:2602.12162v3 Announce Type: replace Abstract: In structurally constrained molecular optimization, state-of-the-art methods restart an expensive oracle-driven search from scratch for every new in

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

Amortized Multi-Objective Optimization Across Tasks with Generative Solution Modeling

DGX agent

arXiv:2511.09598v5 Announce Type: replace Abstract: Many real-world applications require solving families of expensive multi-objective optimization problems~(EMOPs) under varying operational condition

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

An Embarrassingly Simple Graph Heuristic Reveals Shortcut-Solvable Benchmarks for Sequential Recommendation

DGX agent

arXiv:2605.07125v1 Announce Type: cross Abstract: Sequential recommendation has increasingly shifted toward generative recommenders that combine sequential patterns with semantic item information. Yet

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

An Interpretable and Scalable Framework for Evaluating Large Language Models

DGX agent

arXiv:2605.07046v1 Announce Type: cross Abstract: Evaluation of large language models (LLMs) is increasingly critical, yet standard benchmarking methods rely on average accuracy, overlooking both the

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Anatomy of Unlearning: The Dual Impact of Fact Salience and Model Fine-Tuning

DGX agent

arXiv:2602.19612v3 Announce Type: replace Abstract: Machine Unlearning (MU) enables Large Language Models (LLMs) to remove unsafe or outdated information. However, existing work assumes that all facts

model-releasesarxiv-cs-cl
11 May 2026
Model Releases

Ask Early, Ask Late, Ask Right: When Does Clarification Timing Matter for Long-Horizon Agents?

DGX agent

arXiv:2605.07937v1 Announce Type: new Abstract: Long-horizon AI agents execute complex workflows spanning hundreds of sequential actions, yet a single wrong assumption early on can cascade into irreve

model-releasesarxiv-cs-cl
11 May 2026
Model Releases

Ask Patients with Patience: Enabling LLMs for Human-Centric Medical Dialogue with Grounded Reasoning

DGX agent

arXiv:2502.07143v3 Announce Type: replace Abstract: The severe shortage of medical doctors limits access to timely and reliable healthcare, leaving millions underserved. Large language models (LLMs) o

model-releasesarxiv-cs-cl
11 May 2026
Model Releases

Attention Transfer Is Not Universally Effective for Vision Transformers

DGX agent

arXiv:2605.07191v1 Announce Type: new Abstract: A recent work shows that Attention Transfer, which transfers only the attention patterns from a pre-trained teacher Vision Transformer (ViT) to a random

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

Attribution-Based Neuron Utility for Plasticity Restoration in Deep Networks

DGX agent

arXiv:2605.06834v1 Announce Type: new Abstract: Continual learning research attempts to conserve two fundamental capabilities: new knowledge acquisition and the preservation of previously acquired kno

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

Bayesian Fine-tuning in Projected Subspaces

DGX agent

arXiv:2605.07706v1 Announce Type: new Abstract: Low-Rank Adaptation (LoRA) enables parameter-efficient fine-tuning of large models by decomposing weight updates into low-rank matrices, significantly r

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

Behavior Cue Reasoning: Monitorable Reasoning Improves Efficiency and Safety through Oversight

DGX agent

arXiv:2605.07021v1 Announce Type: new Abstract: Reasoning in Large Language Models (LLMs) poses a challenge for oversight as many misaligned behaviors do not surface until reasoning concludes. To addr

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Benchmarked Yet Not Measured -- Generative AI Should be Evaluated Against Real-World Utility

DGX agent

arXiv:2605.06856v1 Announce Type: cross Abstract: Generative AI systems achieve impressive performance on standard benchmarks yet fail to deliver real-world utility, a disconnect we identify across 28

model-releasesarxiv-cs-cl
11 May 2026
Model Releases

Benchmarking EngGPT2-16B-A3B against Comparable Italian and International Open-source LLMs

DGX agent

arXiv:2605.07731v1 Announce Type: cross Abstract: This report benchmarks the performance of ENGINEERING Ingegneria Informatica S.p.A.'s EngGPT2MoE-16B-A3B LLM, a 16B parameter Mixture of Experts (MoE)

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Benchmarking Foundation Models for Renal Lesion Stratification in CT

DGX agent

arXiv:2605.07749v1 Announce Type: new Abstract: The rapid proliferation of open-source medical foundation models (FMs) raises a practical question: how well do their pre-trained representations transf

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

Benchmarking World-Model Learning with Environment-Level Queries

DGX agent

arXiv:2510.19788v4 Announce Type: replace Abstract: World models are central to building AI agents capable of flexible reasoning and planning. Yet current evaluations (i) test only properties measurab

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Beyond Factor Aggregation: Gauge-Aware Low-Rank Server Representations for Federated LoRA

DGX agent

arXiv:2605.06733v1 Announce Type: cross Abstract: Federated LoRA enables parameter-efficient adaptation of large language models under decentralized data and limited client resources.However, directly

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Beyond Factual Accuracy: Evaluating Global Reasoning Integrity in RAG Systems with LogicScore

DGX agent

arXiv:2601.15050v4 Announce Type: replace Abstract: Current evaluation methods for Retrieval Augmented Generation (RAG) suffer from extit{factual myopia}: they relentlessly emphasize factual accuracy

model-releasesarxiv-cs-cl
11 May 2026
Model Releases

Beyond GSD-as-Token: Continuous Scale Conditioning for Remote Sensing VLMs

DGX agent

arXiv:2605.07562v1 Announce Type: new Abstract: Remote sensing vision-language models (RS-VLMs) face a fundamental mismatch with natural-image counterparts: the same geographic object exhibits radical

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

Beyond Linear Attention: Softmax Transformers Implement In-Context Reinforcement Learning

DGX agent

arXiv:2605.07333v1 Announce Type: new Abstract: In-context reinforcement learning (ICRL) studies agents that, after pretraining, adapt to new tasks by conditioning on additional context without parame

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

Beyond LoRA vs. Full Fine-Tuning: Gradient-Guided Optimizer Routing for LLM Adaptation

DGX agent

arXiv:2605.07111v1 Announce Type: cross Abstract: Recent literature on fine-tuning Large Language Models highlights a fundamental debate. While Full Fine-Tuning (FFT) provides the representational pla

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Beyond Retrieval: A Multitask Benchmark and Model for Code Search

DGX agent

arXiv:2605.04615v2 Announce Type: replace-cross Abstract: Code search has usually been evaluated as first-stage retrieval, even though production systems rely on broader pipelines with reranking and d

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Beyond the Black Box: Interpretability of Agentic AI Tool Use

DGX agent

arXiv:2605.06890v1 Announce Type: new Abstract: AI agents are promising for high-stakes enterprise workflows, but dependable deployment remains limited because tool-use failures are difficult to diagn

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

BGM-IV: an AI-powered Bayesian generative modeling approach for instrumental variable analysis

DGX agent

arXiv:2605.07029v1 Announce Type: cross Abstract: Instrumental-variable (IV) regression enables causal estimation under endogeneity, but modern IV problems often involve nonlinear structural effects a

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Bi3: A Biplatform, Bicultural, Biperson Dataset for Social Robot Navigation

DGX agent

arXiv:2605.06863v1 Announce Type: new Abstract: We contribute Bi3, a dataset of social robot navigation among groups of people in a constrained lab space. Compared to prior data collection efforts for

model-releasesarxiv-cs-ro
11 May 2026
Model Releases

BioProVLA-Agent: An Affordable, Protocol-Driven, Vision-Enhanced VLA-Enabled Embodied Multi-Agent System with Closed-Loop-Capable Reasoning for Biological Laboratory Manipulation

DGX agent

arXiv:2605.07306v1 Announce Type: cross Abstract: Biological laboratory automation can reduce repetitive manual work and improve reproducibility, but reliable embodied execution in wet-lab environment

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

bispectrum: Selective G-Bispectra Made Practical

DGX agent

arXiv:2605.07270v1 Announce Type: new Abstract: Many machine learning tasks are invariant under the action of a group G of transformations: signal classification can be invariant under translations, i

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

BoHA: Blockwise Hadamard Product Adaptation for Parameter-Efficient Fine-Tuning

DGX agent

arXiv:2509.21637v2 Announce Type: replace Abstract: Parameter-efficient fine-tuning (PEFT) of large language models trains a small task-specific parameter set while keeping the pretrained model frozen

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

Breaking Spatial Uniformity: Prior-Guided Mamba with Radial Serialization for Lens Flare Removal

DGX agent

arXiv:2605.07650v1 Announce Type: new Abstract: Lens flares, caused by complex optical aberrations, severely degrade image quality especially in nighttime photography. Although recent restoration meth

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

BRIDGE: Background Routing and Isolated Discrete Gating for Coarse-Mask Local Editing

DGX agent

arXiv:2605.07846v1 Announce Type: new Abstract: Coarse-mask local image editing asks a model to modify a user-indicated region while preserving the surrounding scene. In practice, however, rough masks

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

Bridging the Last Mile of Circuit Design: PostEDA-Bench, a Hierarchical Benchmark for PPA Convergence and DRC Fixing

DGX agent

arXiv:2605.06936v1 Announce Type: cross Abstract: LLM-based agents are increasingly applied to the 'last mile' of Electronic Design Automation (EDA): repairing residual sign-off Design Rule Check (DRC

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

CA-SQL: Complexity-Aware Inference Time Reasoning for Text-to-SQL via Exploration and Compute Budget Allocation

DGX agent

arXiv:2605.08057v1 Announce Type: cross Abstract: While recent advancements in inference-time learning have improved LLM reasoning on Text-to-SQL tasks, current solutions still struggle to perform wel

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Can Agents Price a Reaction? Evaluating LLMs on Chemical Cost Reasoning

DGX agent

arXiv:2605.07251v1 Announce Type: new Abstract: Large Language Models (LLMs) have become increasingly capable as tool-using agents, with benchmarks spanning diverse general agentic tasks. Yet rigorous

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Can You Break RLVER? Probing Adversarial Robustness of RL-Trained Empathetic Agents

DGX agent

arXiv:2605.07138v1 Announce Type: new Abstract: Reinforcement learning from verifiable emotion rewards RLVER has produced language models with strong empathetic performance, evaluated on benchmarks th

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

CarCrashNet: A Large-Scale Dataset and Hierarchical Neural Solver for Data-Driven Structural Crash Simulation

DGX agent

arXiv:2605.07098v1 Announce Type: new Abstract: Crash simulation is a cornerstone of modern vehicle development because it reduces the need for costly physical prototypes, accelerates safety-driven de

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

Chain-based Distillation for Effective Initialization of Variable-Sized Small Language Models

DGX agent

arXiv:2605.07783v1 Announce Type: new Abstract: Large language models (LLMs) achieve strong performance but remain costly to deploy in resource-constrained settings. Training small language models (SL

model-releasesarxiv-cs-cl
11 May 2026
Model Releases

ChartREG++: Towards Benchmarking and Improving Chart Referring Expression Grounding under Diverse referring clues and Multi-Target Referring

DGX agent

arXiv:2605.07415v1 Announce Type: cross Abstract: Referring expression grounding is a core problem in visual grounding and is widely used as a diagnostic of spatial grounding and reasoning in vision a

model-releasesarxiv-cs-cl
11 May 2026
Model Releases

Clinically Aware Synthetic Image Generation for Concept Coverage in Chest X-ray Models

DGX agent

arXiv:2603.15525v2 Announce Type: replace Abstract: Deep learning models for chest X-ray diagnosis are constrained by limited coverage of clinically meaningful concept combinations in publicly availab

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

CoCoReviewBench: A Completeness- and Correctness-Oriented Benchmark for AI Reviewers

DGX agent

arXiv:2605.07905v1 Announce Type: cross Abstract: Despite the rapid development of AI reviewers, evaluating such systems remains challenging: metrics favor overlap with human reviews over correctness.

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

CommandSwarm: Safety-Aware Natural Language-to-Behavior-Tree Generation for Robotic Swarms

DGX agent

arXiv:2605.07764v1 Announce Type: new Abstract: Natural-language interfaces can make swarm robotics more accessible to non-expert operators, but they must translate ambiguous user intent into executab

model-releasesarxiv-cs-ro
11 May 2026
Model Releases

Continually Evolving Skill Knowledge in Vision Language Action Model

DGX agent

arXiv:2511.18085v4 Announce Type: replace-cross Abstract: Vision-language-action (VLA) models show promising knowledge accumulation ability from pretraining, yet continual learning in VLA remains chal

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Contrast-X: A Multi-Modal Contrast Image Synthesis Benchmark and Universal Modality Flow Matching

DGX agent

arXiv:2601.15884v2 Announce Type: replace Abstract: Contrast-enhanced imaging is central to oncologic diagnosis, but contrast agents can be contraindicated for many of the patients who need them most.

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

Convergence and Emergence of In-Context Reinforcement Learning with Chain of Thought

DGX agent

arXiv:2605.07123v1 Announce Type: new Abstract: In-context reinforcement learning (ICRL) refers to the ability of RL agents to adapt to new tasks at inference time without parameter updates by conditi

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

Convex Optimization with Nested Evolving Feasible Sets

DGX agent

arXiv:2605.07386v1 Announce Type: new Abstract: Convex Optimization with Nested Evolving Feasible Sets (CONES)} is considered where the objective function f remains fixed but the feasible region evolv

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

CrossCult-KIBench: A Benchmark for Cross-Cultural Knowledge Insertion in MLLMs

DGX agent

arXiv:2605.06115v2 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs), trained primarily on English-centric data, frequently generate culturally inappropriate or misaligned resp

model-releasesarxiv-cs-ai
11 May 2026
← Previous
1…263264265266267…361
Next →