AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,288 results
Model Releases

Embedding-based In-Context Prompt Training for Enhancing LLMs as Text Encoders

DGX agent

arXiv:2605.01372v1 Announce Type: new Abstract: Large language models (LLMs) have been widely explored for embedding generation. While recent studies show that in-context learning (ICL) effectively en

model-releasesarxiv-cs-cl
5 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

EmoMM: Benchmarking and Steering MLLM for Multimodal Emotion Recognition under Conflict and Missingness

DGX agent

arXiv:2605.01024v1 Announce Type: new Abstract: Multimodal Emotion Recognition (MER) is critical for interpreting real-world interactions. While Multimodal Large Language Models (MLLM) have shown prom

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

Empowering Heterogeneous Graph Foundation Models via Decoupled Relation Alignment

DGX agent

arXiv:2605.00731v1 Announce Type: cross Abstract: While Graph Foundation Models (GFMs) have achieved remarkable success in homogeneous graphs, extending them to multi-domain heterogeneous graphs (MDHG

model-releasesarxiv-cs-ai
5 May 2026
Model Releases

ENFORCE: Nonlinear Constrained Learning with Adaptive-depth Neural Projection

DGX agent

arXiv:2502.06774v4 Announce Type: replace Abstract: Ensuring neural networks adhere to domain-specific constraints is crucial for addressing safety and trustworthiness while also enhancing inference a

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

Enhanced LLM Reasoning by Optimizing Reward Functions with Search-Driven Reinforcement Learning

DGX agent

arXiv:2605.02073v1 Announce Type: new Abstract: Mathematical reasoning is a key benchmark for large language models. Reinforcement learning is a standard post-training mechanism for improving the reas

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

Enhancing Judgment Document Generation via Agentic Legal Information Collection and Rubric-Guided Optimization

DGX agent

arXiv:2605.02011v1 Announce Type: new Abstract: Automating the drafting of judgment documents is pivotal to judicial efficiency, yet it remains challenging due to the dual requirements of comprehensiv

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

Environment-Aware Indoor LoRaWAN Ranging Using Path Loss Model Inversion and Adaptive RSSI Filtering

DGX agent

arXiv:2505.01185v3 Announce Type: replace-cross Abstract: Achieving sub-10 m indoor ranging with LoRaWAN is challenging because multipath, human blockage, and micro-climate dynamics induce non-station

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

ESARBench: A Benchmark for Agentic UAV Embodied Search and Rescue

DGX agent

arXiv:2605.01371v1 Announce Type: new Abstract: The rapid advancement of Multimodal Large Language Models (MLLMs) has empowered Unmanned Aerial Vehicle (UAV) with exceptional capabilities in spatial r

model-releasesarxiv-cs-ro
5 May 2026
Model Releases

Evaluating LLMs on Large-Scale Graph Property Estimation via Random Walks

DGX agent

arXiv:2605.01484v1 Announce Type: new Abstract: With the rapidly improving reasoning abilities of Large Language Models (LLMs), there is also a rising demand to use them in a wide variety of domains.

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

Evaluating Tabular Representation Learning for Network Intrusion Detection

DGX agent

arXiv:2605.02519v1 Announce Type: new Abstract: Classic Network Intrusion Detection Systems (NIDS) often rely on manual feature engineering to extract meaningful patterns from network traffic data. Ho

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

Evolving Token Communication with Parametric Memory Network

DGX agent

arXiv:2605.01869v1 Announce Type: cross Abstract: Token communication has emerged as a promising framework for efficient wireless transmission by representing source data as compact semantic tokens. H

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

Exact Higher-Order Derivatives for SE(3) via Analytical/AD Methods

DGX agent

arXiv:2605.02252v1 Announce Type: new Abstract: Fast prototyping of new SE(3) estimation objectives remains awkward in practice. Modern Lie-group frameworks -- GTSAM, manif, Sophus, SymForce, Ceres --

model-releasesarxiv-cs-ro
5 May 2026
Model Releases

Extracting memorized pieces of (copyrighted) books from open-weight language models

DGX agent

arXiv:2505.12546v5 Announce Type: replace Abstract: Plaintiffs and defendants in copyright lawsuits over generative AI often make sweeping, opposing claims about the extent to which large language mod

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

Extreme Weather Bench: A framework and benchmark for evaluation of high-impact weather

DGX agent

arXiv:2605.01126v1 Announce Type: new Abstract: Forecasting the wide variety of high-impact weather events experienced globally is a challenge for both Artificial Intelligence (AI) and Numerical Weath

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

Federated Reinforcement Learning for Efficient Mobile Crowdsensing under Incomplete Information

DGX agent

arXiv:2605.02705v1 Announce Type: new Abstract: Mobile crowdsensing (MCS) is a distributed sensing architecture that utilizes existing sensors on mobile units (MUs) to perform sensing tasks. A mobile

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

FedPLT: Scalable, Resource-Efficient, and Heterogeneity-Aware Federated Learning via Partial Layer Training

DGX agent

arXiv:2605.02337v1 Announce Type: cross Abstract: Federated Learning (FL) has gained significant attention in distributed machine learning by enabling collaborative model training across decentralized

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

Feedback-Normalized Developer Memory for Reinforcement-Learning Coding Agents: A Safety-Gated MCP Architecture

DGX agent

arXiv:2605.01567v1 Announce Type: cross Abstract: Large language model (LLM) coding agents increasingly operate over repositories, terminals, tests, and execution traces across long software-engineeri

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

FeedbackLLM: Metadata driven Multi-Agentic Language Agnostic Test Case Generator with Evolving prompt and Coverage Feedback

DGX agent

arXiv:2605.01264v1 Announce Type: cross Abstract: Traditional approaches to test case generation often involve manual effort and incur significant computational overhead. Additionally, these approache

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

Fin-PRM: A Domain-Specialized Process Reward Model for Financial Reasoning in Large Language Models

DGX agent

arXiv:2508.15202v2 Announce Type: replace Abstract: Process Reward Models (PRMs) supervise intermediate reasoning steps in large language models (LLMs), but existing PRMs are mainly trained on general

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

Fine-Tuning Impairs the Balancedness of Foundation Models in Long-tailed Personalized Federated Learning

DGX agent

arXiv:2605.02247v1 Announce Type: new Abstract: Personalized federated learning (PFL) with foundation models has emerged as a promising paradigm enabling clients to adapt to heterogeneous data distrib

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

Finite-Sample Analysis of Elimination in Active Hypothesis Testing

DGX agent

arXiv:2605.01039v1 Announce Type: new Abstract: A fixed-confidence, finite-sample problem of active hypothesis testing arises in many safety-critical applications. Situated in the context of sequentia

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

Flexi-LoRA with Input-Adaptive Ranks: Efficient Finetuning for Speech and Reasoning Tasks

DGX agent

arXiv:2605.01959v1 Announce Type: cross Abstract: Parameter-efficient fine-tuning methods like Low-Rank Adaptation (LoRA) have become essential for deploying large language models, yet their static pa

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

FlexSQL: Flexible Exploration and Execution Make Better Text-to-SQL Agents

DGX agent

arXiv:2605.02815v1 Announce Type: new Abstract: Text-to-SQL over large analytical databases requires navigating complex schemas, resolving ambiguous queries, and grounding decisions in actual data. Mo

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

Focus on the Core: Empowering Diffusion Large Language Models by Self-Contrast

DGX agent

arXiv:2605.01373v1 Announce Type: new Abstract: The iterative denoising paradigm of Diffusion Large Language Models (DLMs) endows them with a distinct advantage in global context modeling. However, cu

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

FoR-Net: Learning to Focus on Hard Regions for Efficient Semantic Segmentation

DGX agent

arXiv:2605.02764v1 Announce Type: new Abstract: We present FoR-Net, a lightweight architecture for semantic segmentation that focuses on identifying and enhancing hard regions. Instead of relying on h

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

From Euler to Dormand-Prince: ODE Solvers for Flow Matching Generative Models

DGX agent

arXiv:2605.00836v1 Announce Type: new Abstract: Sampling from Flow Matching generative models requires solving an ordinary differential equation (ODE) whose computational cost is dominated by neural n

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

From Where Things Are to What They Are For: Benchmarking Spatial-Functional Intelligence in Multimodal LLMs

DGX agent

arXiv:2605.02130v1 Announce Type: new Abstract: Human-level agentic intelligence extends beyond low-level geometric perception, evolving from recognizing where things are to understanding what they ar

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

FT-RAG: A Fine-grained Retrieval-Augmented Generation Framework for Complex Table Reasoning

DGX agent

arXiv:2605.01495v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) enhances Large Language Models (LLMs) by grounding responses in external knowledge during inference. However, conve

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

G-reasoner: Foundation Models for Unified Reasoning over Graph-structured Knowledge

DGX agent

arXiv:2509.24276v4 Announce Type: replace Abstract: Large language models (LLMs) excel at complex reasoning but remain limited by static and incomplete parametric knowledge. Retrieval-augmented genera

model-releasesarxiv-cs-ai
5 May 2026
Model Releases

GameScope: A Multi-Attribute, Multi-Codec Benchmark Dataset for Gaming Video Quality Assessment

DGX agent

arXiv:2605.01272v1 Announce Type: new Abstract: The development of video game streaming has grown rapidly, with major platforms such as YouTube and Twitch using different codecs. To support quality as

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

GaMMA: Towards Joint Global-Temporal Music Understanding in Large Multimodal Models

DGX agent

arXiv:2605.00371v1 Announce Type: cross Abstract: In this paper, we propose GaMMA, a state-of-the-art (SoTA) large multimodal model (LMM) designed to achieve comprehensive musical content understandin

model-releasesarxiv-cs-ai
5 May 2026
Model Releases

Gated Relational Alignment via Confidence-based Distillation for Efficient VLMs

DGX agent

arXiv:2601.22709v3 Announce Type: replace Abstract: Vision-Language Models (VLMs) achieve strong multimodal performance but are costly to deploy, and post-training quantization often causes significan

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

GAZE: Grounded Agentic Zero-shot Evaluation with Viewer-Level Tools and Literature Retrieval on Rare Brain MRI

DGX agent

arXiv:2605.00876v1 Announce Type: cross Abstract: Vision-language models (VLMs) read an image and produce text in a single forward pass, whereas radiologists typically inspect an image several times a

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

GD-FPS: Growth-Driven Feedforward Parameter Selection for Efficient Fine-Tuning

DGX agent

arXiv:2510.27359v2 Announce Type: replace Abstract: Parameter-Efficient Fine-Tuning (PEFT) has emerged as a key strategy for adapting large-scale pre-trained models to downstream tasks, but existing a

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

Gen-Searcher: Reinforcing Agentic Search for Image Generation

DGX agent

arXiv:2603.28767v2 Announce Type: replace Abstract: Recent image generation models have shown strong capabilities in generating high-fidelity and photorealistic images. However, they are fundamentally

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

GeoContra: From Fluent GIS Code to Verifiable Spatial Analysis with Geography-Grounded Repair

DGX agent

arXiv:2605.00782v1 Announce Type: cross Abstract: Reliable spatial analysis in GIScience requires preserving coordinate semantics, topology, units, and geographic plausibility. Current LLM-based GIS s

model-releasesarxiv-cs-ai
5 May 2026
Model Releases

Geospatial foundation-model embeddings improve population estimation unevenly across space and scale

DGX agent

arXiv:2605.01650v1 Announce Type: new Abstract: Reliable subnational population estimates are essential for applications, yet remain difficult where censuses are sparse, outdated or spatially coarse.

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

GR-Ben: A General Reasoning Benchmark for Evaluating Process Reward Models

DGX agent

arXiv:2605.01203v1 Announce Type: cross Abstract: Currently, process reward models (PRMs) have exhibited remarkable potential for test-time scaling. Since large language models (LLMs) regularly genera

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

Gradient Boosting within a Single Attention Layer

DGX agent

arXiv:2604.03190v2 Announce Type: replace Abstract: Transformer attention computes a single softmax-weighted average over values -- a one-pass estimate that cannot correct its own errors. We introduce

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

GraphLand: Evaluating Graph Machine Learning Models on Diverse Industrial Data

DGX agent

arXiv:2409.14500v5 Announce Type: replace Abstract: Although data that can be naturally represented as graphs is widespread in real-world applications across diverse industries, popular graph ML bench

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

Grounding Synthetic Data Generation With Vision and Language Models

DGX agent

arXiv:2603.09625v2 Announce Type: replace Abstract: Deep learning models benefit from increasing data diversity and volume, motivating synthetic data augmentation to improve existing datasets. However

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

Growing Transformers: Modular Composition and Layer-wise Expansion on a Frozen Substrate

DGX agent

arXiv:2507.07129v3 Announce Type: replace-cross Abstract: We study a constrained training regime for decoder-only Transformers in which the token interface is fixed, previously trained dense blocks ar

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

HalluScan: A Systematic Benchmark for Detecting and Mitigating Hallucinations in Instruction-Following LLMs

DGX agent

arXiv:2605.02443v1 Announce Type: new Abstract: Large Language Models (LLMs) have demonstrated remarkable capabilities across diverse natural language processing tasks, yet they remain susceptible to

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

HARMES: A Multi-Modal Dataset for Wearable Human Activity Recognition with Motion, Environmental Sensing and Sound

DGX agent

arXiv:2605.02596v1 Announce Type: new Abstract: With each sensing modality exhibiting inherent strengths and limitations, multi-modal approaches for wearable Human Activity Recognition (HAR) are becom

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

How Reasoning Evolves from Post-Training Data: An Empirical Study Using Chess

DGX agent

arXiv:2604.05134v2 Announce Type: replace Abstract: We study how reasoning evolves in a language model -- from supervised fine-tuning (SFT) to reinforcement learning (RL) -- by analyzing how a set of

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

How Well Can We Decode Vowels from Auditory EEG -- A Rigorous Cross-Subject Benchmark with Honest Assessment

DGX agent

arXiv:2605.00865v1 Announce Type: cross Abstract: EEG based phoneme decoding is promising for brain computer interfaces, but many prior studies rely on within subject evaluation, small cohorts, or wea

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

Human Cognitive Benchmarks Reveal Foundational Visual Gaps in MLLMs

DGX agent

arXiv:2502.16435v4 Announce Type: replace-cross Abstract: Humans develop perception through a bottom-up hierarchy: from basic primitives and Gestalt principles to high-level semantics. In contrast, cu

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

Implicit Neural Representation-Based Continuous Single Image Super-Resolution: An Empirical Benchmark

DGX agent

arXiv:2601.17723v2 Announce Type: replace Abstract: Implicit neural representation (INR) has become the standard approach for arbitrary-scale image super-resolution (ASSR). To date, no empirical study

model-releasesarxiv-cs-cv
5 May 2026
← Previous
1…278279280281282…361
Next →