AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
16,965 results
Model Releases

Many Ways to Be Fake: Benchmarking Fake News Detection Under Strategy-Driven AI Generation

DGX agent

arXiv:2604.09514v1 Announce Type: new Abstract: Recent advances in large language models (LLMs) have enabled the large-scale generation of highly fluent and deceptive news-like content. While prior wo

model-releasesarxiv-cs-cl
13 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

MARINER: A 3E-Driven Benchmark for Fine-Grained Perception and Complex Reasoning in Open-Water Environments

DGX agent

arXiv:2604.08615v1 Announce Type: cross Abstract: Fine-grained visual understanding and high-level reasoning in real-world open-water environments remain under-explored due to the lack of dedicated be

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

MATCHA: Efficient Deployment of Deep Neural Networks on Multi-Accelerator Heterogeneous Edge SoCs

DGX agent

arXiv:2604.09124v1 Announce Type: cross Abstract: Deploying DNNs on System-on-Chips (SoC) with multiple heterogeneous acceleration engines is challenging, and the majority of deployment frameworks can

model-releasesarxiv-cs-lg
13 Apr 2026
Model Releases

MedConceal: A Benchmark for Clinical Hidden-Concern Reasoning Under Partial Observability

DGX agent

arXiv:2604.08788v1 Announce Type: new Abstract: Patient-clinician communication is an asymmetric-information problem: patients often do not disclose fears, misconceptions, or practical barriers unless

model-releasesarxiv-cs-cl
13 Apr 2026
Model Releases

Medical Reasoning with Large Language Models: A Survey and MR-Bench

DGX agent

arXiv:2604.08559v1 Announce Type: cross Abstract: Large language models (LLMs) have achieved strong performance on medical exam-style tasks, motivating growing interest in their deployment in real-wor

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

Memory-efficient Continual Learning with Prototypical Exemplar Condensation

DGX agent

arXiv:2603.13804v2 Announce Type: replace-cross Abstract: Rehearsal-based continual learning (CL) mitigates catastrophic forgetting by maintaining a subset of samples from previous tasks for replay. E

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

Memory-Efficient Transfer Learning with Fading Side Networks via Masked Dual Path Distillation

DGX agent

arXiv:2604.09088v1 Announce Type: new Abstract: Memory-efficient transfer learning (METL) approaches have recently achieved promising performance in adapting pre-trained models to downstream tasks. Th

model-releasesarxiv-cs-cv
13 Apr 2026
Model Releases

Mitigating Extrinsic Gender Bias for Bangla Classification Tasks

DGX agent

arXiv:2411.10636v2 Announce Type: replace-cross Abstract: In this study, we investigate extrinsic gender bias in Bangla pretrained language models, a largely underexplored area in low-resource languag

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

Mnemis: Dual-Route Retrieval on Hierarchical Graphs for Long-Term LLM Memory

DGX agent

arXiv:2602.15313v2 Announce Type: replace Abstract: AI Memory, specifically how models organizes and retrieves historical messages, becomes increasingly valuable to Large Language Models (LLMs), yet e

model-releasesarxiv-cs-cl
13 Apr 2026
Model Releases

MolPaQ: Modular Quantum-Classical Patch Learning for Interpretable Molecular Generation

DGX agent

arXiv:2604.08575v1 Announce Type: cross Abstract: Molecular generative models must jointly ensure validity, diversity, and property control, yet existing approaches typically trade off among these obj

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

MONETA: Multimodal Industry Classification through Geographic Information with Multi Agent Systems

DGX agent

arXiv:2604.07956v2 Announce Type: replace Abstract: Industry classification schemes are integral parts of public and corporate databases as they classify businesses based on economic activity. Due to

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

Multivariate Time Series Anomaly Detection via Dual-Branch Reconstruction and Autoregressive Flow-based Residual Density Estimation

DGX agent

arXiv:2604.08582v1 Announce Type: cross Abstract: Multivariate Time Series Anomaly Detection (MTSAD) is critical for real-world monitoring scenarios such as industrial control and aerospace systems. M

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

Natural Riemannian gradient for learning functional tensor networks

DGX agent

arXiv:2604.09263v1 Announce Type: cross Abstract: We consider machine learning tasks with low-rank functional tree tensor networks (TTN) as the learning model. While in the case of least-squares regre

model-releasesarxiv-cs-lg
13 Apr 2026
Model Releases

NCL-BU at SemEval-2026 Task 3: Fine-tuning XLM-RoBERTa for Multilingual Dimensional Sentiment Regression

DGX agent

arXiv:2604.08923v1 Announce Type: new Abstract: Dimensional Aspect-Based Sentiment Analysis (DimABSA) extends traditional ABSA from categorical polarity labels to continuous valence-arousal (VA) regre

model-releasesarxiv-cs-cl
13 Apr 2026
Model Releases

Neural networks for Text-to-Speech evaluation

DGX agent

arXiv:2604.08562v1 Announce Type: cross Abstract: Ensuring that Text-to-Speech (TTS) systems deliver human-perceived quality at scale is a central challenge for modern speech technologies. Human subje

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

Nexus: Same Pretraining Loss, Better Downstream Generalization via Common Minima

DGX agent

arXiv:2604.09258v1 Announce Type: new Abstract: Pretraining is the cornerstone of Large Language Models (LLMs), dominating the vast majority of computational budget and data to serve as the primary en

model-releasesarxiv-cs-lg
13 Apr 2026
Model Releases

Noise-Aware In-Context Learning for Hallucination Mitigation in ALLMs

DGX agent

arXiv:2604.09021v1 Announce Type: cross Abstract: Auditory large language models (ALLMs) have demonstrated strong general capabilities in audio understanding and reasoning tasks. However, their reliab

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

NTIRE 2026 The 3rd Restore Any Image Model (RAIM) Challenge: Multi-Exposure Image Fusion in Dynamic Scenes (Track 2)

DGX agent

arXiv:2604.09030v1 Announce Type: new Abstract: This paper presents NTIRE 2026, the 3rd Restore Any Image Model (RAIM) challenge on multi-exposure image fusion in dynamic scenes. We introduce a benchm

model-releasesarxiv-cs-cv
13 Apr 2026
Model Releases

On Semiotic-Grounded Interpretive Evaluation of Generative Art

DGX agent

arXiv:2604.08641v1 Announce Type: cross Abstract: Interpretation is essential to deciphering the language of art: audiences communicate with artists by recovering meaning from visual artifacts. Howeve

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

On-the-Fly Adaptation to Quantization: Configuration-Aware LoRA for Efficient Fine-Tuning of Quantized LLMs

DGX agent

arXiv:2509.25214v3 Announce Type: replace-cross Abstract: As increasingly large pre-trained models are released, deploying them on edge devices for privacy-preserving applications requires effective c

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

Online Intention Prediction via Control-Informed Learning

DGX agent

arXiv:2604.09303v1 Announce Type: cross Abstract: This paper presents an online intention prediction framework for estimating the goal state of autonomous systems in real time, even when intention is

model-releasesarxiv-cs-lg
13 Apr 2026
Model Releases

PACED: Distillation and On-Policy Self-Distillation at the Frontier of Student Competence

DGX agent

arXiv:2603.11178v3 Announce Type: replace Abstract: Standard LLM distillation treats all training problems equally -- wasting compute on problems the student has already mastered or cannot yet solve.

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

Parameterized Complexity Of Representing Models Of MSO Formulas

DGX agent

arXiv:2604.08707v1 Announce Type: new Abstract: Monadic second order logic (MSO2) plays an important role in parameterized complexity due to the Courcelle's theorem. This theorem states that the probl

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

PhysInOne: Visual Physics Learning and Reasoning in One Suite

DGX agent

arXiv:2604.09415v1 Announce Type: cross Abstract: We present PhysInOne, a large-scale synthetic dataset addressing the critical scarcity of physically-grounded training data for AI systems. Unlike exi

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

PilotBench: A Benchmark for General Aviation Agents with Safety Constraints

DGX agent

arXiv:2604.08987v1 Announce Type: new Abstract: As Large Language Models (LLMs) advance toward embodied AI agents operating in physical environments, a fundamental question emerges: can models trained

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

PinpointQA: A Dataset and Benchmark for Small Object-Centric Spatial Understanding in Indoor Videos

DGX agent

arXiv:2604.08991v1 Announce Type: cross Abstract: Small object-centric spatial understanding in indoor videos remains a significant challenge for multimodal large language models (MLLMs), despite its

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

Precise Shield: Explaining and Aligning VLLM Safety via Neuron-Level Guidance

DGX agent

arXiv:2604.08881v1 Announce Type: new Abstract: In real-world deployments, Vision-Language Large Models (VLLMs) face critical challenges from multilingual and multimodal composite attacks: harmful ima

model-releasesarxiv-cs-cv
13 Apr 2026
Model Releases

Precomputing Multi-Agent Path Replanning using Temporal Flexibility

DGX agent

arXiv:2601.04884v2 Announce Type: replace Abstract: Executing a multi-agent plan can be challenging when an agent is delayed, because this typically creates conflicts with other agents. So, we need to

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

Pretrain-then-Adapt: Uncertainty-Aware Test-Time Adaptation for Text-based Person Search

DGX agent

arXiv:2604.08598v1 Announce Type: cross Abstract: Text-based person search faces inherent limitations due to data scarcity, driven by stringent privacy constraints and the high cost of manual annotati

model-releasesarxiv-cs-cv
13 Apr 2026
Model Releases

Provable Post-Training Quantization: Theoretical Analysis of OPTQ and Qronos

DGX agent

arXiv:2508.04853v2 Announce Type: replace-cross Abstract: Post-training quantization (PTQ) has become a crucial tool for reducing the memory and compute costs of modern deep neural networks, including

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

QARIMA: A Quantum Approach To Classical Time Series Analysis

DGX agent

arXiv:2604.08277v2 Announce Type: replace-cross Abstract: We present a quantum-inspired ARIMA methodology that integrates quantum-assisted lag discovery with fixed-configuration variational quantum ci

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

QoS-QoE Translation with Large Language Model

DGX agent

arXiv:2604.08703v1 Announce Type: cross Abstract: QoS-QoE translation is a fundamental problem in multimedia systems because it characterizes how measurable system and network conditions affect user-p

model-releasesarxiv-cs-lg
13 Apr 2026
Model Releases

QuanBench+: A Unified Multi-Framework Benchmark for LLM-Based Quantum Code Generation

DGX agent

arXiv:2604.08570v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used for code generation, yet quantum code generation is still evaluated mostly within single frameworks

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

Quantisation Reshapes the Metacognitive Geometry of Language Models

DGX agent

arXiv:2604.08976v1 Announce Type: new Abstract: We report that model quantisation restructures domain-level metacognitive efficiency in LLMs rather than degrading it uniformly. Evaluating Llama-3-8B-I

model-releasesarxiv-cs-cl
13 Apr 2026
Model Releases

R2G: A Multi-View Circuit Graph Benchmark Suite from RTL to GDSII

DGX agent

arXiv:2604.08810v1 Announce Type: new Abstract: Graph neural networks (GNNs) are increasingly applied to physical design tasks such as congestion prediction and wirelength estimation, yet progress is

model-releasesarxiv-cs-cv
13 Apr 2026
Model Releases

RADSeg: Unleashing Parameter and Compute Efficient Zero-Shot Open-Vocabulary Segmentation Using Agglomerative Models

DGX agent

arXiv:2511.19704v2 Announce Type: replace Abstract: Open-vocabulary semantic segmentation (OVSS) underpins many vision and robotics tasks that require generalizable semantic understanding. Existing ap

model-releasesarxiv-cs-cv
13 Apr 2026
Model Releases

RansomTrack: A Hybrid Behavioral Analysis Framework for Ransomware Detection

DGX agent

arXiv:2604.08739v1 Announce Type: cross Abstract: Ransomware poses a serious and fast-acting threat to critical systems, often encrypting files within seconds of execution. Research indicates that ran

model-releasesarxiv-cs-lg
13 Apr 2026
Model Releases

Reasoning in a Combinatorial and Constrained World: Benchmarking LLMs on Natural-Language Combinatorial Optimization

DGX agent

arXiv:2602.02188v2 Announce Type: replace Abstract: While large language models (LLMs) have shown strong performance in math and logic reasoning, their ability to handle combinatorial optimization (CO

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

ReplicatorBench: Benchmarking LLM Agents for Replicability in Social and Behavioral Sciences

DGX agent

arXiv:2602.11354v2 Announce Type: replace Abstract: The literature has witnessed an emerging interest in AI agents for automated assessment of scientific papers. Existing benchmarks focus primarily on

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

RESample: A Robust Data Augmentation Framework via Exploratory Sampling for Robotic Manipulation

DGX agent

arXiv:2510.17640v3 Announce Type: replace-cross Abstract: Vision-Language-Action (VLA) models have demonstrated remarkable performance on complex tasks through imitation learning in recent robotic man

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

Retrieval Augmented Classification for Confidential Documents

DGX agent

arXiv:2604.08628v1 Announce Type: cross Abstract: Unauthorized disclosure of confidential documents demands robust, low-leakage classification. In real work environments, there is a lot of inflow and

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

Revisiting Image Manipulation Localization under Realistic Manipulation Scenarios

DGX agent

arXiv:2509.20006v3 Announce Type: replace Abstract: With the large models easing the labor-intensive manipulation process, image manipulations in today's real scenarios often entail a complex manipula

model-releasesarxiv-cs-cv
13 Apr 2026
Model Releases

Robust Reasoning Benchmark

DGX agent

arXiv:2604.08571v1 Announce Type: cross Abstract: While Large Language Models (LLMs) achieve high performance on standard mathematical benchmarks, their underlying reasoning processes remain highly ov

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

SafeAdapt: Provably Safe Policy Updates in Deep Reinforcement Learning

DGX agent

arXiv:2604.09452v1 Announce Type: cross Abstract: Safety guarantees are a prerequisite to the deployment of reinforcement learning (RL) agents in safety-critical tasks. Often, deployment environments

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

SAGE: A Service Agent Graph-guided Evaluation Benchmark

DGX agent

arXiv:2604.09285v1 Announce Type: new Abstract: The development of Large Language Models (LLMs) has catalyzed automation in customer service, yet benchmarking their performance remains challenging. Ex

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

SEA-Eval: A Benchmark for Evaluating Self-Evolving Agents Beyond Episodic Assessment

DGX agent

arXiv:2604.08988v1 Announce Type: new Abstract: Current LLM-based agents demonstrate strong performance in episodic task execution but remain constrained by static toolsets and episodic amnesia, faili

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

See, Hear, and Understand: Benchmarking Audiovisual Human Speech Understanding in Multimodal Large Language Models

DGX agent

arXiv:2512.02231v2 Announce Type: replace-cross Abstract: Multimodal large language models (MLLMs) are expected to jointly interpret vision, audio, and language, yet existing video benchmarks rarely a

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

Seeing is Believing: Robust Vision-Guided Cross-Modal Prompt Learning under Label Noise

DGX agent

arXiv:2604.09532v1 Announce Type: cross Abstract: Prompt learning is a parameter-efficient approach for vision-language models, yet its robustness under label noise is less investigated. Visual conten

model-releasesarxiv-cs-ai
13 Apr 2026
← Previous
1…344345346347348…354
Next →