AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,288 results
Model Releases

Drift is a Sampling Error: SNR-Aware Power Distributions for Long-Horizon Robotic Planning

DGX agent

arXiv:2605.09537v1 Announce Type: new Abstract: Despite rapid progress in Vision-Language-Action (VLA) models for robotic control, instruction drift remains a persistent failure mode in long-horizon t

model-releasesarxiv-cs-ro
12 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

DRNet: All-in-One Image Restoration via Prior-Guided Dynamic Reparameterization

DGX agent

arXiv:2605.08627v1 Announce Type: new Abstract: All-in-one image restoration aims to handle diverse degradations within a single model. However, existing methods often suffer from three key limitation

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

DSGBench: A Diverse Strategic Game Benchmark for Evaluating LLM-based Agents in Complex Decision-Making Environments

DGX agent

arXiv:2503.06047v2 Announce Type: replace Abstract: Large language model (LLM)-based agents are increasingly applied to complex strategic environments that demand long-horizon reasoning, multi-agent i

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

DUET: Optimize Token-Budget Allocation for Reinforcement Learning with Verifiable Rewards

DGX agent

arXiv:2605.08441v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) generates hundreds of thousands of tokens per training step, with rollout generation dominating

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Dynamics-Aligned Shared Hypernetworks for Contextual RL under Discontinuous Shifts

DGX agent

arXiv:2602.06550v2 Announce Type: replace-cross Abstract: Zero-shot generalization in contextual reinforcement learning remains a core challenge, particularly when the context is latent and must be in

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Echo-LoRA: Parameter-Efficient Fine-Tuning via Cross-Layer Representation Injection

DGX agent

arXiv:2605.08177v1 Announce Type: cross Abstract: Parameter-efficient fine-tuning (PEFT) has become a practical route for adapting large language models to downstream tasks, with LoRA-style methods be

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

EchoAlign: Bridging Generative and Discriminative Learning under Noisy Labels

DGX agent

arXiv:2405.12969v3 Announce Type: replace Abstract: Noisy labels severely hinder the accuracy and generalization of machine learning models, especially when ambiguous instance features make reliable a

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

EcoGym: Evaluating LLMs for Long-Horizon Plan-and-Execute in Interactive Economies

DGX agent

arXiv:2602.09514v3 Announce Type: replace-cross Abstract: Long-horizon planning is widely recognized as a core capability of autonomous LLM-based agents; however, current evaluation frameworks suffer

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

EconWebArena: Benchmarking Autonomous Agents on Economic Tasks in Realistic Web Environments

DGX agent

arXiv:2506.08136v3 Announce Type: replace Abstract: We introduce EconWebArena, a benchmark for evaluating autonomous agents on complex, multimodal economic tasks in realistic web environments. The ben

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

Edge-specific signal propagation on mature chromophore-region 3D mechanism graphs for fluorescent protein quantum-yield prediction

DGX agent

arXiv:2605.06644v2 Announce Type: replace Abstract: Fluorescent protein quantum yield (QY) is governed by the mature chromophore and its three-dimensional microenvironment rather than sequence identit

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

EdgeFlowerTune: Evaluating Federated LLM Fine-Tuning Under Realistic Edge System Constraints

DGX agent

arXiv:2605.08636v1 Announce Type: new Abstract: Federated fine-tuning offers a promising paradigm for adapting large language models (LLMs) on edge devices by leveraging the rich, diverse, and continu

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

EduStory: A Unified Framework for Pedagogically-Consistent Multi-Shot STEM Instructional Video Generation

DGX agent

arXiv:2605.09378v1 Announce Type: cross Abstract: Long-horizon video generation has advanced in visual quality, yet existing methods still struggle to maintain knowledge consistency and coherent pedag

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Efficient Ensemble Selection from Binary and Pairwise Feedback

DGX agent

arXiv:2605.09588v1 Announce Type: cross Abstract: Organizations increasingly deploy multiple AI systems across task domains, but selecting a small, high-performing ensemble can require costly model ca

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Efficient Evaluation of LLM Performance with Statistical Guarantees

DGX agent

arXiv:2601.20251v3 Announce Type: replace-cross Abstract: Exhaustively evaluating many large language models (LLMs) on a large suite of benchmarks is expensive. We cast benchmarking as finite-populati

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Efficient Neural Architectures for Real-Time ECG Interpretation on Limited Hardware

DGX agent

arXiv:2605.09848v1 Announce Type: new Abstract: Electrocardiogram (ECG) interpretation is essential for diagnosing a wide range of cardiac abnormalities. While deep learning has shown strong potential

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

EgoMemReason: A Memory-Driven Reasoning Benchmark for Long-Horizon Egocentric Video Understanding

DGX agent

arXiv:2605.09874v1 Announce Type: cross Abstract: Next-generation visual assistants, such as smart glasses, embodied agents, and always-on life-logging systems, must reason over an entire day or more

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

EmbodiSkill: Skill-Aware Reflection for Self-Evolving Embodied Agents

DGX agent

arXiv:2605.10332v1 Announce Type: new Abstract: Embodied agents can benefit from skills that guide object search, action execution, and state changes across diverse environments. Since embodied enviro

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

EmoS: A High-Fidelity Multimodal Benchmark for Fine-grained Streaming Emotional Understanding

DGX agent

arXiv:2605.08847v1 Announce Type: new Abstract: In the context of today's high-pressure, aging society, the demand for large-scale emotional models capable of providing empathetic support is more crit

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

EnactToM: An Evolving Benchmark for Functional Theory of Mind in Embodied Agents

DGX agent

arXiv:2605.09826v1 Announce Type: new Abstract: Theory of Mind (ToM), the ability to track others epistemic state, makes humans efficient collaborators. AI agents need the same capacity in multi agent

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

EnergyLens: Interpretable Closed-Form Energy Models for Multimodal LLM Inference Serving

DGX agent

arXiv:2605.10556v1 Announce Type: new Abstract: As large language models span dense, mixture-of-experts, and state-space architectures and are deployed on heterogeneous accelerators under increasingly

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

EpiGraph: A Knowledge Graph and Benchmark for Evidence-Intensive Reasoning in Epilepsy

DGX agent

arXiv:2605.09505v1 Announce Type: new Abstract: Epilepsy diagnosis and treatment require evidence-intensive reasoning across heterogeneous clinical knowledge, including biosignal patterns, genetic mec

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Equilibrium Residuals Expose Three Regimes of Matrix-Game Strategic Reasoning in Language Models

DGX agent

arXiv:2605.10410v1 Announce Type: new Abstract: Large language models can score well on named game-theory benchmarks while failing on the same strategic computation once semantic cues are removed. We

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

ER-Reason: A Benchmark Dataset for LLM Clinical Reasoning in the Emergency Room

DGX agent

arXiv:2505.22919v3 Announce Type: replace Abstract: Existing benchmarks for evaluating the clinical reasoning capabilities of large language models (LLMs) often lack a clear definition of 'clinical re

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

ERIS: Enhancing Privacy and Scalability in Federated Learning via Federated Shard Aggregation

DGX agent

arXiv:2602.08617v2 Announce Type: replace Abstract: Scaling Federated Learning (FL) to billion-parameter models forces a challenging trade-off between privacy, scalability, and model utility. Existing

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

EverydayMMQA: A Multilingual and Multimodal Framework for Culturally Grounded Spoken Visual QA

DGX agent

arXiv:2510.06371v2 Announce Type: replace-cross Abstract: Large-scale multimodal models achieve strong results on tasks like Visual Question Answering (VQA), but they are often limited when queries re

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Exactness Matters for Physical Rule Enforcement

DGX agent

arXiv:2605.08285v1 Announce Type: new Abstract: Autoregressive scientific forecasters often enforce physical or structural constraints by repairing each predicted state before feeding it back into the

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Explanation Fairness in Large Language Models: An Empirical Analysis of Disparities in How LLMs Justify Decisions Across Demographic Groups

DGX agent

arXiv:2605.08671v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly deployed not only to make decisions but to explain them. While AI decision fairness has been studied ext

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Explicit Reasoning Makes Better Judges: A Systematic Study on Accuracy, Efficiency, and Robustness

DGX agent

arXiv:2509.13332v2 Announce Type: replace Abstract: As Large Language Models (LLMs) are increasingly adopted as automated judges in benchmarking and reward modeling, ensuring their reliability, effici

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Exploitation Without Deception: Dark Triad Feature Steering Reveals Separable Antisocial Circuits in Language Models

DGX agent

arXiv:2605.09773v1 Announce Type: cross Abstract: We use sparse autoencoder (SAE) feature steering to amplify Dark Triad personality traits (Machiavellianism, narcissism, and psychopathy) in Llama-3.3

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Exploring the AI Obedience: Why is Generating a Pure Color Image Harder than CyberPunk?

DGX agent

arXiv:2603.00166v2 Announce Type: replace-cross Abstract: Recent advances in generative AI have shown human-level performance in complex content creation. However, we identify a 'Paradox of Simplicity

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

expo: Exploration-prioritized policy optimization via adaptive kl regulation and gaussian curriculum sampling

DGX agent

arXiv:2605.09923v1 Announce Type: new Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has become the standard paradigm for LLM mathematical reasoning, where Group Relative Policy Optim

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

FactoryNet: A Large-Scale Dataset toward Industrial Time-Series Foundation Models

DGX agent

arXiv:2605.09081v1 Announce Type: cross Abstract: We introduce the first universal pretraining corpus for industrial time-series data: FactoryNet. 51M datapoints across 23k end-to-end task executions

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Failing Forward: Adaptive Failure-Informed Learning for Vision-Language-Action Models

DGX agent

arXiv:2605.08434v1 Announce Type: new Abstract: Vision-language-action (VLA) models provide a promising paradigm for scalable robotic manipulation, yet their reliance on success-only behavioral clonin

model-releasesarxiv-cs-ro
12 May 2026
Model Releases

Fashion Florence: Fine-Tuning Florence-2 for Structured Fashion Attribute Extraction

DGX agent

arXiv:2605.09827v1 Announce Type: cross Abstract: We present Fashion Florence, a Florence-2 vision-language model fine-tuned with LoRA to extract structured fashion attributes from clothing images. Gi

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Fashion130K: An E-commerce Fashion Dataset for Outfit Generation with Unified Multi-modal Condition

DGX agent

arXiv:2605.10127v1 Announce Type: new Abstract: Recent research work on fashion outfit generation focuses on promoting visual consistency of garments by leveraging key information from reference image

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

Feature Repulsion and Spectral Lock-in: An Empirical Study of Two-Layer Network Grokking

DGX agent

arXiv:2605.08119v1 Announce Type: cross Abstract: Tian (2025) proves a repulsion theorem (Theorem 6) for the matrix B = (widetilde{F}^op widetilde{F} + eta I)^{-1} during the interactive feature-learn

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Feature Rivalry in Sparse Autoencoder Representations: A Mechanistic Study of Uncertainty-Driven Feature Competition in LLMs

DGX agent

arXiv:2605.08149v1 Announce Type: cross Abstract: Sparse Autoencoders (SAEs) decompose large language model representations into interpretable features, but how these features interact under uncertain

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

Featurized Occupation Measures for Structured Global Search in Numerical Optimal Control

DGX agent

arXiv:2603.16231v2 Announce Type: replace-cross Abstract: Numerical optimal control has long been split between globally structured but dimensionally intractable Hamilton--Jacobi--Bellman (HJB) method

model-releasesarxiv-cs-ro
12 May 2026
Model Releases

Federated Language Models Under Bandwidth Budgets: Distillation Rates and Conformal Coverage

DGX agent

arXiv:2605.09986v1 Announce Type: cross Abstract: Training a language model on data scattered across bandwidth-limited nodes that cannot be centralized is a setting that arises in clinical networks, e

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

Filtering Memorization from Parameter-Space in Diffusion Models

DGX agent

arXiv:2605.10439v1 Announce Type: new Abstract: Low-Rank Adaptation (LoRA) has become a widely used mechanism for customizing diffusion models, enabling users to inject new visual concepts or styles t

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

Fin-Bias: Comprehensive Evaluation for LLM Decision-Making under human bias in Finance Domain

DGX agent

arXiv:2605.09106v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed in financial contexts, raising critical concerns about reliability, alignment, and susceptibility

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

FinTSB: A Comprehensive and Practical Benchmark for Financial Time Series Forecasting

DGX agent

arXiv:2502.18834v2 Announce Type: replace-cross Abstract: Financial time series (FinTS) record the behavior of human-brain-augmented decision-making, capturing valuable historical information that can

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Fitting Multilinear Polynomials for Logic Gate Networks

DGX agent

arXiv:2605.08657v1 Announce Type: cross Abstract: We study learnable logic gate networks that stack layers of 2-input Boolean gates to build combinational circuits. Every 2-input gate has a unique mul

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Fix the Loss, Not the Radius: Rethinking the Adversarial Perturbation of Sharpness-Aware Minimization

DGX agent

arXiv:2605.10183v1 Announce Type: new Abstract: Sharpness-Aware Minimization (SAM) improves generalization by minimizing the worst-case loss within a fixed parameter-space radius neighborhood. SAM and

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

FLAME: Adaptive Mixture-of-Experts for Continual Multimodal Multi-Task Learning

DGX agent

arXiv:2605.09355v1 Announce Type: new Abstract: Real-world model deployment across multiple domains requires multimodal models to operate under two complementary regimes: (1) multi-task pretraining, t

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Flame3D: Zero-shot Compositional Reasoning of 3D Scenes with Agentic Language Models

DGX agent

arXiv:2605.09218v1 Announce Type: cross Abstract: 3D scene understanding spans reasoning about free space, object grounding, hypothetical object insertions, complex geometric relationships, and integr

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

FlashClear: Ultra-Fast Image Content Removal via Efficient Step Distillation and Feature Caching

DGX agent

arXiv:2605.09003v1 Announce Type: new Abstract: Recently, diffusion-based object removal models have achieved impressive results in eliminating objects and their associated visual effects. However, th

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

Follow the Mean: Reference-Guided Flow Matching

DGX agent

arXiv:2605.10302v1 Announce Type: new Abstract: Existing approaches to controllable generation typically rely on fine-tuning, auxiliary networks, or test-time search. We show that flow matching admits

model-releasesarxiv-cs-lg
12 May 2026
← Previous
1…253254255256257…361
Next →