AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,288 results
Model Releases

Generalized Euler Logarithm and its Applications in Machine Learning: Natural Gradient, Backpropagation, Generalized EG, Mirror Descent and OLPS

DGX agent

arXiv:2502.17500v3 Announce Type: replace-cross Abstract: This paper investigates in depth the fundamental properties of the two-parameter generalized Euler logarithm and its inverse, the associated d

model-releasesarxiv-cs-ai
11 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

GLiGuard: Schema-Conditioned Classification for LLM Safeguard

DGX agent

arXiv:2605.07982v1 Announce Type: new Abstract: Ensuring safe, policy-compliant outputs from large language models requires real-time content moderation that can scale across multiple safety dimension

model-releasesarxiv-cs-cl
11 May 2026
Model Releases

Globally Optimal Training of Spiking Neural Networks via Parameter Reconstruction

DGX agent

arXiv:2605.08022v1 Announce Type: cross Abstract: Spiking Neural Networks (SNNs) have been proposed as biologically plausible and energy-efficient alternatives to conventional Artificial Neural Networ

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Goal-Conditioned Decision Transformer for Multi-Goal Offline Reinforcement Learning

DGX agent

arXiv:2410.06347v2 Announce Type: replace-cross Abstract: Reinforcement learning (RL) in robotics faces significant hurdles regarding sample efficiency and generalization across varying goals. While O

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Gradient-Based LoRA Rank Allocation Under GRPO: An Empirical Study

DGX agent

arXiv:2605.07366v1 Announce Type: new Abstract: Adaptive rank allocation for LoRA, allocating more parameters to important layers and fewer to unimportant ones, consistently improves efficiency under

model-releasesarxiv-cs-cl
11 May 2026
Model Releases

Gradient Extrapolation-Based Policy Optimization

DGX agent

arXiv:2605.06755v1 Announce Type: cross Abstract: Reinforcement learning is widely used to improve the reasoning ability of large language models, especially when answers can be automatically checked.

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Gradient Starvation in Binary-Reward GRPO: Why Group-Mean Centering Fails and Why the Simplest Fix Works

DGX agent

arXiv:2605.07689v1 Announce Type: new Abstract: Group Relative Policy Optimization (GRPO) is a standard algorithm for reinforcement learning from verifiable rewards, but its group-mean-centered advant

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

Graph Representation Learning Augmented Model Manipulation on Federated Fine-Tuning of LLMs

DGX agent

arXiv:2605.07961v1 Announce Type: new Abstract: Federated fine-tuning (FFT) has emerged as a privacy-preserving paradigm for collaboratively adapting large language models (LLMs). Built upon federated

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

Graph-Structured Hyperdimensional Computing for Data-Efficient and Explainable Process-Structure-Property Prediction

DGX agent

arXiv:2605.07999v1 Announce Type: cross Abstract: Multiphoton photoreduction enables high-fidelity fabrication of complex 3D microstructures, yet reliable process-structure-property (PSP) prediction r

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

GraphReAct: Reasoning and Acting for Multi-step Graph Inference

DGX agent

arXiv:2605.07357v1 Announce Type: new Abstract: Reasoning-acting frameworks enhance large language models (LLMs) by interleaving reasoning with actions for dynamic information acquisition. However, ex

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

GSM-SEM: Benchmark and Framework for Generating Semantically Variant Augmentations

DGX agent

arXiv:2605.07053v1 Announce Type: cross Abstract: Benchmarks like GSM8K are popular measures of mathematical reasoning, but leaderboard gains can overstate true capability due to memorization of fixed

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Hallucination Detection via Activations of Open-Weight Proxy Analyzers

DGX agent

arXiv:2605.07209v1 Announce Type: cross Abstract: We introduce a proxy-analyzer framework for detecting hallucinations in large language models. Instead of looking inside the generating model, our sys

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Have Graph -- Will Lift? The Case for Higher-Order Benchmarks

DGX agent

arXiv:2605.07397v1 Announce Type: new Abstract: After a somewhat rocky start, geometry and topology have established a foothold in machine learning. Message passing, either on graphs or higher-order c

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

Head Similarity: Modeling Structured Whole-Head Appearance Beyond Face Recognition

DGX agent

arXiv:2605.07766v1 Announce Type: new Abstract: Many vision applications require identity consistency beyond strict biometric recognition, especially under non-frontal views or when facial cues are mi

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

Hierarchical Dual-Subspace Decoupling for Continual Learning in Vision-Language Models

DGX agent

arXiv:2605.07512v1 Announce Type: new Abstract: Class-incremental learning aims to continuously acquire new knowledge while preserving previously learned information, thereby mitigating catastrophic f

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

Hierarchical Task Network Planning with LLM-Generated Heuristics

DGX agent

arXiv:2605.07707v1 Announce Type: new Abstract: HTN planning is a variation of classical planning where, instead of searching for a linear sequence of actions, an algorithm decomposes higher-level tas

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

How Big Should a Wireless Foundation Model Be?

DGX agent

arXiv:2605.07266v1 Announce Type: cross Abstract: Wireless foundation models are rapidly emerging as a key enabler of AI-native communication systems, yet a fundamental question remains unanswered: ho

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

How Far Are VLMs from Privacy Awareness in the Physical World? An Empirical Study

DGX agent

arXiv:2605.05340v2 Announce Type: replace-cross Abstract: As Vision-Language Models (VLMs) are increasingly deployed as autonomous cognitive cores for embodied assistants, evaluating their privacy awa

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

How Far Is Document Parsing from Solved? PureDocBench: A Source-TraceableBenchmark across Clean, Degraded, and Real-World Settings

DGX agent

arXiv:2605.07492v1 Announce Type: new Abstract: The past year has seen over 20 open-source document parsing models, yet thefield still benchmarks almost exclusively on OmniDocBench, a 1,355-pagemanual

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

HumanNet: Scaling Human-centric Video Learning to One Million Hours

DGX agent

arXiv:2605.06747v1 Announce Type: new Abstract: Progress in embodied intelligence increasingly depends on scalable data infrastructure. While vision and language have scaled with internet corpora, lea

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

HyperEyes: Dual-Grained Efficiency-Aware Reinforcement Learning for Parallel Multimodal Search Agents

DGX agent

arXiv:2605.07177v1 Announce Type: cross Abstract: Existing multimodal search agents process target entities sequentially, issuing one tool call per entity and accumulating redundant interaction rounds

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Implicit Preference Alignment for Human Image Animation

DGX agent

arXiv:2605.07545v1 Announce Type: cross Abstract: Human image animation has witnessed significant advancements, yet generating high-fidelity hand motions remains a persistent challenge due to their hi

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

In-Context Credit Assignment via the Core

DGX agent

arXiv:2605.06920v1 Announce Type: cross Abstract: We propose incentive-aligned mechanisms for in-context credit assignment: the task of assigning credit for AI-generated content (e.g. code, news artic

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Inference Time Causal Probing in LLMs

DGX agent

arXiv:2605.07631v1 Announce Type: new Abstract: Causal probing methods aim to test and control how internal representations influence the behavior of generative models. In causal probing, an intervent

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Intent-Driven Semantic ID Generation for Grounded Conversational News Recommendation

DGX agent

arXiv:2605.07613v1 Announce Type: new Abstract: Conversational news recommendation requires grounding each suggestion in a rapidly evolving article corpus while addressing implicit user intents that l

model-releasesarxiv-cs-cl
11 May 2026
Model Releases

IntentGrasp: A Comprehensive Benchmark for Intent Understanding

DGX agent

arXiv:2605.06832v1 Announce Type: cross Abstract: Accurately understanding the intent behind speech, conversation, and writing is crucial to the development of helpful Large Language Model (LLM) assis

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

InterLV-Search: Benchmarking Interleaved Multimodal Agentic Search

DGX agent

arXiv:2605.07510v1 Announce Type: cross Abstract: Existing benchmarks for multimodal agentic search evaluate multimodal search and visual browsing, but visual evidence is either confined to the input

model-releasesarxiv-cs-cl
11 May 2026
Model Releases

Interpreting Reinforcement Learning Agents with Susceptibilities

DGX agent

arXiv:2605.08007v1 Announce Type: new Abstract: Susceptibilities are a technique for neural network interpretability that studies the response of posterior expectation values of observables to perturb

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

Is Your Prompt Poisoning Code? Defect Induction Rates and Security Mitigation Strategies

DGX agent

arXiv:2510.22944v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have become indispensable for automated code generation, yet the quality and security of their outputs remain a c

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Knowing but Not Correcting: Routine Task Requests Suppress Factual Correction in LLMs

DGX agent

arXiv:2605.05957v2 Announce Type: replace Abstract: LLMs reliably correct false claims when presented in isolation, yet when the same claims are embedded in task-oriented requests, they often comply r

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

LARAG: Link-Aware Retrieval Strategy for RAG Systems in Hyperlinked Technical Documentation

DGX agent

arXiv:2605.07517v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) enhances the factual grounding of Large Language Models by conditioning their outputs on external documents. Howe

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Learning Agent Routing From Early Experience

DGX agent

arXiv:2605.07180v1 Announce Type: new Abstract: LLM agents achieve strong performance on complex reasoning tasks but incur high latency and compute cost. In practice, many queries fall within the capa

model-releasesarxiv-cs-cl
11 May 2026
Model Releases

Learning and Reusing Policy Decompositions for Hierarchical Generalized Planning with LLM Agents

DGX agent

arXiv:2605.06957v1 Announce Type: new Abstract: We present a dynamic policy-learning approach that combines generalized planning and hierarchical task decomposition for LLM-based agents. Our method, H

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Learning Material-Aware Hamiltonian Risk Fields for Safe Navigation

DGX agent

arXiv:2605.07038v1 Announce Type: new Abstract: Risk-aware navigation should be selective: a policy should expose evasive degrees of freedom only when the local scene admits a lower-risk feasible mane

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

LithoBench: Benchmarking Large Multimodal Models for Remote-Sensing Lithology Interpretation

DGX agent

arXiv:2605.07640v1 Announce Type: cross Abstract: Remote sensing lithology interpretation is fundamental to geological surveys, mineral exploration, and regional geological mapping. Unlike general lan

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

LLM-Based Agents for Competitive Landscape Mapping in Drug Asset Due Diligence

DGX agent

arXiv:2508.16571v4 Announce Type: replace Abstract: In this paper, we describe and benchmark a competitor-discovery component used within an agentic AI system for fast drug asset due diligence. A comp

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Mage: Multi-Axis Evaluation of LLM-Generated Executable Game Scenes Beyond Compile-Pass Rate

DGX agent

arXiv:2605.07342v1 Announce Type: cross Abstract: Compile-pass rate is the dominant evaluation signal for LLM code generation, yet for multi-component domain-specific artifacts it can be actively misl

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

MAS-Algorithm: A Workflow for Solving Algorithmic Programming Problems with a Multi-Agent System

DGX agent

arXiv:2605.05949v2 Announce Type: replace Abstract: Algorithmic problem solving serves as a rigorous testbed for evaluating structured reasoning in AI coding systems, as it directly reflects a model's

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Mask2Cause: Causal Discovery via Adjacency Constrained Causal Attention

DGX agent

arXiv:2605.07280v1 Announce Type: cross Abstract: Leveraging deep learning for causal discovery in time series remains challenging because existing neural methods predominantly rely on component-wise

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Mathematical Reasoning via Intervention-Based Time-Series Causal Discovery Using LLMs as Concept Mastery Simulators

DGX agent

arXiv:2605.07600v1 Announce Type: cross Abstract: Recent methods for improving LLM mathematical reasoning, whether through MCTS-based test-time search or causal graph-guided knowledge injection, canno

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

MathlibPR: Pull Request Merge-Readiness Benchmark for Formal Mathematical Libraries

DGX agent

arXiv:2605.07147v1 Announce Type: cross Abstract: The ecosystem of Lean and Mathlib has become the de facto standard for large language model (LLM) assisted formal reasoning with remarkable successes

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

MatryoshkaLoRA: Learning Accurate Hierarchical Low-Rank Representations for LLM Fine-Tuning

DGX agent

arXiv:2605.07850v1 Announce Type: cross Abstract: With the rise in scale for deep learning models to billions of parameters, the computational cost of fine-tuning remains a significant barrier to depl

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

MAVEN: Multi-Agent Verification-Elaboration Network with In-Step Epistemic Auditing

DGX agent

arXiv:2605.07646v1 Announce Type: cross Abstract: While explicit reasoning trajectories enhance model interpretability, existing paradigms often rely on monolithic chains that lack intermediate verifi

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

McNdroid: A Longitudinal Multimodal Benchmark for Robust Drift Detection in Android Malware

DGX agent

arXiv:2605.06894v1 Announce Type: cross Abstract: Machine learning (ML) in real-world systems must contend with concept drift, adversarial actors, and a spectrum of potential features with varying cos

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

Mean-Pooled Cosine Similarity is Not Length-Invariant: Theory and Cross-Domain Evidence for a Length-Invariant Alternative

DGX agent

arXiv:2605.07345v1 Announce Type: new Abstract: Mean-pooled cosine similarity is the default metric for comparing neural representations across languages, modalities, and tasks. We establish that this

model-releasesarxiv-cs-cl
11 May 2026
Model Releases

MedAction: Towards Active Multi-turn Clinical Diagnostic LLMs

DGX agent

arXiv:2605.07305v1 Announce Type: cross Abstract: Most existing LLM diagnoses are evaluated on static, single-turn settings where complete patient information is provided upfront, an oversimplificatio

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

MedVIGIL: Evaluating Trustworthy Medical VLMs Under Broken Visual Evidence

DGX agent

arXiv:2605.07919v1 Announce Type: new Abstract: Medical vision--language models (VLMs) are usually evaluated on intact image--question pairs, but trustworthy clinical use requires a stronger property:

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

MELD: Multi-Task Equilibrated Learning Detector for AI-Generated Text

DGX agent

arXiv:2605.06903v1 Announce Type: cross Abstract: Large language models are now embedded in everyday writing workflows, making reliable AI-generated text detection important for academic integrity, co

model-releasesarxiv-cs-ai
11 May 2026
← Previous
1…265266267268269…361
Next →