AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,606Total entries
1Added by human
84,605Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,288 results
Model Releases

DySink: Dynamic Frame Sinks for Autoregressive Long Video Generation

DGX agent

arXiv:2605.21028v1 Announce Type: new Abstract: Autoregressive long video generation often adopts bounded-memory streaming for efficiency, typically combining local windows for short-term continuity w

model-releasesarxiv-cs-cv
21 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

ECUAS{n}: A family of metrics for principled evaluation of uncertainty-augmented systems

DGX agent

arXiv:2605.20490v1 Announce Type: cross Abstract: In high-stakes automated decision-making, access to predictive uncertainty is essential for enabling users -- human or downstream systems -- to accept

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

Enhanced Reinforcement Learning-based Process Synthesis via Quantum Computing

DGX agent

arXiv:2605.21213v1 Announce Type: cross Abstract: In this work, we present quantum reinforcement learning (RL) as a solution strategy for process synthesis problems. Building on our prior work, we dev

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

Evolutionary Generation of Multi-Agent Systems

DGX agent

arXiv:2602.06511v3 Announce Type: replace Abstract: Large language model (LLM)-based multi-agent systems (MAS) show strong promise for complex reasoning, planning, and tool-augmented tasks, but design

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

Explainability Methods for Hardware Trojan Detection: A Systematic Comparison

DGX agent

arXiv:2601.18696v4 Announce Type: replace Abstract: Hardware trojans are malicious circuits which compromise the functionality and security of an integrated circuit (IC). These circuits are manufactur

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

Exploring Deep Learning and Ultra-Widefield Imaging for Diabetic Retinopathy and Macular Edema

DGX agent

arXiv:2603.08235v2 Announce Type: replace Abstract: Diabetic retinopathy (DR) and diabetic macular edema (DME) are leading causes of preventable blindness among working-age adults. Traditional approac

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

FAIR-Pruner: A Flexible Framework for Automatic Layer-Wise Pruning via Tolerance of Difference

DGX agent

arXiv:2508.02291v3 Announce Type: replace Abstract: Structured pruning is a standard tool for compressing deep neural networks, but its practical performance depends on how sparsity is allocated acros

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

FBOS-RL: Feedback-Driven Bi-Objective Synergistic Reinforcement Learning

DGX agent

arXiv:2605.20256v1 Announce Type: new Abstract: Reinforcement learning has become a cornerstone for aligning and unlocking the reasoning capabilities of large-scale models. At its core, the training l

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

FedCoE: Bridging Generalization and Personalization via Federated Coordinated Dual-level MoEs

DGX agent

arXiv:2605.21264v1 Announce Type: new Abstract: Federated Learning (FL) has emerged as a promising paradigm for privacy-preserving distributed learning. However, existing FL methods face a fundamental

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

FedCritic: Serverless Federated Critic Learning-based Resource Allocation for Multi-Cell OFDMA in 6G

DGX agent

arXiv:2605.21418v1 Announce Type: cross Abstract: In sixth-generation (6G) ultra-dense networks, aggressive frequency reuse amplifies inter-cell interference (ICI), making multi-cell orthogonal freque

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

Federated LoRA Fine-Tuning for LLMs via Collaborative Alignment

DGX agent

arXiv:2605.21217v1 Announce Type: cross Abstract: Low-rank adaptation (LoRA) has emerged as a powerful tool for parameter-efficient fine-tuning of large language models (LLMs). This paper studies LoRA

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy

DGX agent

arXiv:2605.20965v1 Announce Type: new Abstract: Large Vision-Language Models (LVLMs) have shown remarkable performance on a wide range of vision-language tasks. Despite this progress, they are still p

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

Findings of the Counter Turing Test: AI-Generated Text Detection

DGX agent

arXiv:2605.20761v1 Announce Type: new Abstract: The rapid proliferation of AI-generated text has introduced significant challenges in maintaining the integrity of digital content. Advanced generative

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

Fine-grained Claim-level RAG Benchmark for Law

DGX agent

arXiv:2605.21071v1 Announce Type: new Abstract: The rapid progress of large language models (LLMs) is shifting semantic search toward a question-answering paradigm, where users ask questions and LLMs

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

Free-Grained Hierarchical Visual Recognition

DGX agent

arXiv:2510.14737v3 Announce Type: replace Abstract: Hierarchical image recognition seeks to predict class labels along a semantic taxonomy, from broad categories to specific ones, typically under the

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

FT-Dojo: Towards Autonomous LLM Fine-Tuning with Language Agents

DGX agent

arXiv:2603.01712v2 Announce Type: replace-cross Abstract: Fine-tuning large language models for vertical domains remains labor-intensive, requiring practitioners to curate data, configure training, an

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

FullFlow: Upgrading Text-to-Image Flow Matching Models for Bidirectional Vision--Language Generation

DGX agent

arXiv:2605.20316v1 Announce Type: new Abstract: Modern text-to-image diffusion models encode rich visual priors, but expose them only through one-way text-conditioned generation. Existing unified visi

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

Gated Normalization Removal and Scale Anchoring in Pre-Norm Transformers

DGX agent

arXiv:2602.10408v2 Announce Type: replace-cross Abstract: Normalization layers are standard in transformers, but it is not clear whether their sample-dependent computations are necessary throughout bo

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

GenAI-Driven Threat Detection with Microsoft Security Copilot

DGX agent

arXiv:2605.20896v1 Announce Type: cross Abstract: Defending against today's increasingly sophisticated cyberattacks requires security analysts to continuously translate evolving attacker tradecraft in

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

Geometry-Lite: Interpretable Safety Probing via Layer-Wise Margin Geometry

DGX agent

arXiv:2605.20241v1 Announce Type: cross Abstract: Prompt-level safety probes for large language models use hidden-state representations to separate safe from unsafe prompts, but strong average detecti

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

GradPower: Powering Gradients for Faster Language Model Pre-Training

DGX agent

arXiv:2505.24275v3 Announce Type: replace Abstract: We propose GradPower, a lightweight gradient-transformation technique for accelerating language model pre-training. Given a gradient vector g=(g_i)_

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

GraphRAG on Consumer Hardware: Benchmarking Local LLMs for Healthcare EHR Schema Retrieval

DGX agent

arXiv:2605.20815v1 Announce Type: new Abstract: Graph-based Retrieval Augmented Generation (GraphRAG) extends retrieval-augmented generation to support structured reasoning over complex corpora, but i

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

Group-Algebraic Tensors: Provably-optimal Equivariant Learning and Physical Symmetry Discovery

DGX agent

arXiv:2605.20440v1 Announce Type: new Abstract: We introduce the star_G tensor algebra, in which any finite group G defines the multiplication rule, making equivariance an intrinsic algebraic property

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

Hack-Verifiable Environments: Towards Evaluating Reward Hacking at Scale

DGX agent

arXiv:2605.20744v1 Announce Type: new Abstract: Aligning autonomous agents with human intent remains a central challenge in modern AI. A key manifestation of this challenge is reward hacking, whereby

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

HalluCXR: Benchmarking and Mitigating Hallucinations in Medical Vision-Language Models for Chest Radiograph Interpretation

DGX agent

arXiv:2605.20469v1 Announce Type: new Abstract: Vision-language models (VLMs) are increasingly used for medical image interpretation, yet they frequently hallucinate, generating clinically plausible b

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

How Much Online RL is Enough? Informative Rollouts for Offline Preference Optimization in RLVR

DGX agent

arXiv:2605.21266v1 Announce Type: new Abstract: Reinforcement Learning from Verifiable Rewards (RLVR) has emerged as a powerful paradigm for reasoning in language models, with GRPO as its primary exam

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

HRM-Text: Efficient Pretraining Beyond Scaling

DGX agent

arXiv:2605.20613v1 Announce Type: new Abstract: The current pretraining paradigm for large language models relies on massive compute and internet-scale raw text, creating a significant barrier to foun

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

Humanoid Whole-Body Manipulation via Active Spatial Brain and Generalizable Action Cerebellum

DGX agent

arXiv:2605.21133v1 Announce Type: new Abstract: In this paper, we explore spatial-aware humanoid whole-body manipulation task. Compared with tabletop settings, this task poses two key challenges: 1) S

model-releasesarxiv-cs-ro
21 May 2026
Model Releases

Hyper-V2X: Hypernetworks for Estimating Epistemic and Aleatoric Uncertainty in Cooperative Bird's-Eye-View Semantic Segmentation

DGX agent

arXiv:2605.21309v1 Announce Type: new Abstract: Cooperative perception enabled by Vehicle-to-Everything (V2X) communication enhances autonomous driving safety by creating a unified environmental repre

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

Improving Quantized Model Performance in Qualitative Analysis with Multi-Pass Prompt Verification

DGX agent

arXiv:2605.20193v1 Announce Type: new Abstract: Quantized Large Language Models (LLMs) are used more often in qualitative analysis because they run fast and need fewer computing resources. This study

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

InternBootcamp Technical Report: Boosting LLM Reasoning with Verifiable Task Scaling

DGX agent

arXiv:2508.08636v2 Announce Type: replace Abstract: Large language models (LLMs) have revolutionized artificial intelligence by enabling complex reasoning capabilities. While recent advancements in re

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

JFAA: Technical Report for the EPIC-KITCHENS-100 Action Anticipation Challenge at EgoVis 2026

DGX agent

arXiv:2605.20904v1 Announce Type: new Abstract: We propose JFAA, a JEPA-based Future Action Anticipation method for the EPIC-KITCHENS-100 (EK-100) Action Anticipation task. Inspired by the representat

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

JobArabi: An Arabic Corpus and Analysis of Job Announcements from Social Media

DGX agent

arXiv:2605.20960v1 Announce Type: new Abstract: This paper introduces JobArabi, a large-scale corpus of Arabic job announcements collected from social media between January 2024 and October 2025. The

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

JUDO: A Juxtaposed Domain-Oriented Multimodal Reasoner for Industrial Anomaly QA

DGX agent

arXiv:2605.20284v1 Announce Type: new Abstract: Industrial anomaly detection has been significantly advanced by Large Multimodal Models (LMMs), enabling diverse human instructions beyond detection, pa

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

LAION-C: An Out-of-Distribution Benchmark for Web-Scale Vision Models

DGX agent

arXiv:2506.16950v2 Announce Type: replace Abstract: Out-of-distribution (OOD) robustness is a desired property of computer vision models. Improving model robustness requires high-quality signals from

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

LamPO: A Lambda Style Policy Optimization for Reasoning Language Models

DGX agent

arXiv:2605.21235v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) has become an effective paradigm for improving reasoning language models on tasks such as mathemat

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

Large-Step Training Dynamics of a Two-Factor Linear Transformer Model

DGX agent

arXiv:2605.21292v1 Announce Type: cross Abstract: Gradient-flow analyses show that simplified linear transformers can learn the in-context linear-regression algorithm, but they do not explain the fini

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

Lean Refactor: Multi-Objective Controllable Proof Optimization via Agentic Strategy Search

DGX agent

arXiv:2605.20244v1 Announce Type: cross Abstract: We present Lean Refactor, a plug-and-play retrieval-augmented agentic framework for multi-objective, controllable, and version-robust refactoring of L

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

LEAP: A closed-loop framework for perovskite precursor additive discovery

DGX agent

arXiv:2605.20242v1 Announce Type: new Abstract: Efficient discovery of precursor additives is essential for improving the performance of perovskite solar cells, yet the large chemical space makes conv

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

Learning fMRI activations dictionaries across individual geometries via optimal transport

DGX agent

arXiv:2605.20883v1 Announce Type: new Abstract: Dictionary learning is a powerful tool for creating interpretable representations. When applied to functional magnetic resonance imaging (fMRI) data, th

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

Learning Query-Aware Budget-Tier Routing for Runtime Agent Memory

DGX agent

arXiv:2602.06025v2 Announce Type: replace Abstract: Memory is increasingly central to Large Language Model (LLM) agents operating beyond a single context window, yet most existing systems rely on offl

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

Learning-to-Defer with Expert-Conditional Advice

DGX agent

arXiv:2603.14324v3 Announce Type: replace-cross Abstract: Learning-to-Defer routes each input to the expert that minimizes expected cost, but it assumes that the information available to every expert

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

LER-YOLO: Reliability-Aware Expert Routing for Misaligned RGB-Infrared UAV Detection

DGX agent

arXiv:2605.20667v1 Announce Type: new Abstract: Detecting small unmanned aerial vehicles from RGB-infrared remote-sensing pairs remains challenging due to tiny target scale, cluttered backgrounds, and

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

Leveraging LLMs for Grammar Adaptation: A Study on Metamodel-Grammar Co-Evolution

DGX agent

arXiv:2605.21465v1 Announce Type: new Abstract: In model-driven engineering, metamodel evolution leads to the need to adapt corresponding grammars to maintain consistency, which typically requires ted

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

Leveraging Vision-Language Models to Detect Attention in Educational Videos

DGX agent

arXiv:2605.20211v1 Announce Type: new Abstract: Educational videos are a cornerstone of remote and blended learning. However, learners' fluctuating attention remains a significant barrier to effective

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

Llamas on the Web: Memory-Efficient, Performance-Portable, and Multi-Precision LLM Inference with WebGPU

DGX agent

arXiv:2605.20706v1 Announce Type: cross Abstract: Running language models in the browser presents a unique opportunity to build efficient, private, and portable AI applications, but requires contendin

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

LLMs on the Line: Data Determines Loss-to-Loss Scaling Laws

DGX agent

arXiv:2502.12120v3 Announce Type: replace-cross Abstract: Scaling laws guide the development of large language models (LLMs) by offering estimates for the optimal balance of model size, tokens, and co

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

MagBridge-Battery: A Synthetic Bridge Dataset for Li-ion Magnetometry and State-of-Health Diagnostics

DGX agent

arXiv:2605.20240v1 Announce Type: new Abstract: Battery health diagnostics today rely overwhelmingly on electrochemical signals measured at the cell terminals. A parallel literature has shown that mag

model-releasesarxiv-cs-lg
21 May 2026
← Previous
1…213214215216217…361
Next →