AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,288 results
Model Releases

MicroBi-ConvLSTM: An Ultra-Lightweight Efficient Model for Human Activity Recognition on Resource Constrained Devices

DGX agent

arXiv:2602.06523v2 Announce Type: replace Abstract: Human Activity Recognition (HAR) on resource constrained wearables requires models that balance accuracy against strict memory and computational bud

model-releasesarxiv-cs-cv
11 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

MIND: Monge Inception Distance for Generative Models Evaluation

DGX agent

arXiv:2605.06797v1 Announce Type: new Abstract: We propose the Monge Inception Distance (MIND), a metric for evaluating generative models that addresses key limitations of the widely adopted Frechet I

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

MiniAppBench: Evaluating the Shift from Text to Interactive HTML Responses in LLM-Powered Assistants

DGX agent

arXiv:2603.09652v3 Announce Type: replace Abstract: With the rapid advancement of Large Language Models (LLMs) in code generation, human-AI interaction is evolving from static text responses to dynami

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

MIPIAD: Multilingual Indirect Prompt Injection Attack Defense with Qwen -- TF-IDF Hybrid and Meta-Ensemble Learning

DGX agent

arXiv:2605.07269v1 Announce Type: new Abstract: Indirect prompt injection remains a persistent weakness in retrieval-augmented and tool-using LLM systems, and the problem becomes harder to characteris

model-releasesarxiv-cs-cl
11 May 2026
Model Releases

MISA: Mixture of Indexer Sparse Attention for Long-Context LLM Inference

DGX agent

arXiv:2605.07363v1 Announce Type: cross Abstract: DeepSeek Sparse Attention (DSA) sets the state of the art for fine-grained inference-time sparse attention by introducing a learned token-wise indexer

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Mitigating Cognitive Bias in RLHF by Altering Rationality

DGX agent

arXiv:2605.06895v1 Announce Type: new Abstract: How can we make models robust to even imperfect human feedback? In reinforcement learning from human feedback (RLHF), human preferences over model outpu

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

MobileDev-Bench: A Benchmark for Issue Resolution in Mobile Application Development

DGX agent

arXiv:2603.24946v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have shown strong performance on automated software engineering tasks, yet existing benchmarks focus primarily on

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

Model-Driven Policy Optimization in Differentiable Simulators via Stochastic Exploration

DGX agent

arXiv:2605.07520v1 Announce Type: new Abstract: Differentiable planning enables gradient-based optimization of decision-making problems by leveraging differentiable models of system dynamics. However,

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

ModelLens: Finding the Best for Your Task from Myriads of Models

DGX agent

arXiv:2605.07075v1 Announce Type: new Abstract: The open-source model ecosystem now contains hundreds of thousands of pretrained models, yet picking the best model for a new dataset is increasingly in

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

Modular Lie Algebraic PDE Control of Multibody Flexible Manipulators

DGX agent

arXiv:2605.06709v1 Announce Type: new Abstract: This paper addresses PDE-based control for flexible multibody robotic systems, presenting a subsystem-based framework for serial manipulators with arbit

model-releasesarxiv-cs-ro
11 May 2026
Model Releases

More Thinking, More Bias: Length-Driven Position Bias in Reasoning Models

DGX agent

arXiv:2605.06672v1 Announce Type: new Abstract: Chain-of-thought (CoT) reasoning and reasoning-tuned models such as DeepSeek-R1 are commonly assumed to reduce shallow heuristic biases by thinking care

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Multi-Objective Constraint Inference using Inverse reinforcement learning

DGX agent

arXiv:2605.06951v1 Announce Type: new Abstract: Constraint inference is widely considered essential to align reinforcement learning agents with safety boundaries and operational guidelines by observin

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

MultiSoc-4D: A Benchmark for Diagnosing Instruction-Induced Label Collapse in Closed-Set LLM Annotation of Bengali Social Media

DGX agent

arXiv:2605.06940v1 Announce Type: new Abstract: Annotation automation via Large Language Models (LLMs) is the core approach for scaling NLP datasets; however, LLM behavior with respect to closed-set i

model-releasesarxiv-cs-cl
11 May 2026
Model Releases

Muon Dynamics as a Spectral Wasserstein Flow

DGX agent

arXiv:2604.04891v2 Announce Type: replace-cross Abstract: Gradient normalization stabilizes deep-learning optimization, and spectral normalizations are especially natural for matrix-shaped parameter b

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Narrow Secret Loyalty Dodges Black-Box Audits

DGX agent

arXiv:2605.06846v1 Announce Type: cross Abstract: Recent work identifies secret loyalties as a distinct threat from standard backdoors. A secret loyalty causes a model to covertly advance the interest

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

NCL-UoR at SemEval-2026 Task 5: Embedding-Based Methods, Fine-Tuning, and LLMs for Word Sense Plausibility Rating

DGX agent

arXiv:2603.08256v2 Announce Type: replace Abstract: Word sense plausibility rating requires predicting the human-perceived plausibility of a given word sense on a 1-5 scale in the context of short nar

model-releasesarxiv-cs-cl
11 May 2026
Model Releases

Neural Neural Scaling Laws

DGX agent

arXiv:2601.19831v2 Announce Type: replace-cross Abstract: Neural scaling laws predict how language model performance improves with increased training inputs. While aggregate metrics like validation lo

model-releasesarxiv-cs-cl
11 May 2026
Model Releases

Neural Operators as Efficient Function Interpolators

DGX agent

arXiv:2605.07792v1 Announce Type: cross Abstract: Neural operators (NOs) are designed to learn maps between infinite-dimensional function spaces. We propose a novel reframing of their use. By introduc

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

NPMixer: Hierarchical Neighboring Patch Mixing for Time Series Forecasting

DGX agent

arXiv:2605.07476v1 Announce Type: new Abstract: Multivariate time series forecasting remains a challenge due to the complexity of local temporal dynamics and global dependencies across multiple variab

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

NS-Net: Decoupling CLIP Semantic Information through NULL-Space for Generalizable AI-Generated Image Detection

DGX agent

arXiv:2508.01248v4 Announce Type: replace Abstract: The rapid progress of generative models, such as GANs and diffusion models, has facilitated the creation of highly realistic images, raising growing

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

NSMQ Riddles: A Benchmark of Scientific and Mathematical Riddles for Quizzing Large Language Models

DGX agent

arXiv:2605.07051v1 Announce Type: new Abstract: Large Language Models (LLMs) have shown good performance on various science educational benchmarks, demonstrating their potential for use in science and

model-releasesarxiv-cs-cl
11 May 2026
Model Releases

OmicsLM: A Multimodal Large Language Model for Multi-Sample Omics Reasoning

DGX agent

arXiv:2605.06728v1 Announce Type: cross Abstract: Interpreting transcriptomic data is one of the most common analytical tasks in modern biology. Yet most current models either consume expression profi

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

On the Invariance and Generality of Neural Scaling Laws

DGX agent

arXiv:2605.07546v1 Announce Type: new Abstract: Neural scaling laws establish a predictable relationship between model performance and data or compute, offering crucial guidance for resource allocatio

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

Optimal Experiments for Partial Causal Effect Identification

DGX agent

arXiv:2605.06993v1 Announce Type: new Abstract: Causal queries are often only partially identifiable from observational data, and experiments that could tighten the resulting bounds are typically cost

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Optimizing Language Models for Crosslingual Knowledge Consistency

DGX agent

arXiv:2603.04678v2 Announce Type: replace-cross Abstract: Large language models are known to often exhibit inconsistent knowledge. This is particularly problematic in multilingual scenarios, where mod

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

OrScale: Orthogonalised Optimization with Layer-Wise Trust-Ratio Scaling

DGX agent

arXiv:2605.07815v1 Announce Type: cross Abstract: Muon improves neural-network training by orthogonalizing matrix-valued updates, but it leaves each layer's update magnitude controlled mostly by a glo

model-releasesarxiv-cs-cl
11 May 2026
Model Releases

Outlier Smoothing with Closed-Form Rotations for W4A4 Large Language Model Quantization

DGX agent

arXiv:2511.22316v2 Announce Type: replace Abstract: Large Language Models (LLMs) quantization facilitates deploying LLMs in resource-limited settings, but existing methods that combine incompatible gr

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

PAIR-Former: Budgeted Relational Multi-Instance Learning for Functional miRNA Target Prediction

DGX agent

arXiv:2602.00465v3 Announce Type: replace-cross Abstract: Functional miRNA--mRNA targeting is a large-bag prediction problem where each transcript yields a heavy-tailed pool of candidate target sites

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

PathPainter: Transferring the Generalization Ability of Image Generation Models to Embodied Navigation

DGX agent

arXiv:2605.07496v1 Announce Type: new Abstract: Bird's-eye-view (BEV) images have been widely demonstrated to provide valuable prior information for navigation. Given the global information provided b

model-releasesarxiv-cs-ro
11 May 2026
Model Releases

PerCaM-Health: Personalized Dynamic Causal Graphs for Healthcare Reasoning

DGX agent

arXiv:2605.07267v1 Announce Type: new Abstract: Personalized healthcare decisions require reasoning about how physiological and behavioral variables influence an individual patient over time. Existing

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

PerfCoder: Large Language Models for Interpretable Code Performance Optimization

DGX agent

arXiv:2512.14018v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have achieved remarkable progress in automatic code generation, yet their ability to produce high-performance cod

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

PhySPRING: Structure-Preserving Reduction of Physics-Informed Twins via GNN

DGX agent

arXiv:2605.07687v1 Announce Type: new Abstract: Physics-based digital twins aim to predict the dynamics of real-world objects under interaction, enabling real-to-sim-to-real applications in robotics.

model-releasesarxiv-cs-ro
11 May 2026
Model Releases

PolarVLM: Bridging the Semantic-Physical Gap in Vision-Language Models

DGX agent

arXiv:2605.07574v1 Announce Type: new Abstract: Mainstream vision-language models (VLMs) fundamentally struggle with severe optical ambiguities, such as reflections and transparent objects, due to the

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

Pretraining Induces a Reusable Spectral Basis for Downstream Task Adaptation

DGX agent

arXiv:2605.07302v1 Announce Type: new Abstract: Finetuning pretrained models occurs in a low-dimensional subspace of the full parameter space. Prior work has focused on characterizing this optimizatio

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

PRIMED: Adaptive Modality Suppression for Referring Audio-Visual Segmentation via Biased Competition

DGX agent

arXiv:2605.07154v1 Announce Type: new Abstract: Referring Audio-Visual Segmentation (Ref-AVS) seeks to localize and segment target objects in video frames based on visual, auditory, and textual referr

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

ProactiveMobile: A Comprehensive Benchmark for Boosting Proactive Intelligence on Mobile Devices

DGX agent

arXiv:2602.21858v4 Announce Type: replace Abstract: Multimodal large language models (MLLMs) have made significant progress in mobile agent development, yet their capabilities are predominantly confin

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

ProcObject-10K: Benchmarking Object-Centric Procedural Understanding in Instructional Videos

DGX agent

arXiv:2512.03479v2 Announce Type: replace Abstract: Procedural activities are fundamentally driven by object state transitions, yet existing instructional video benchmarks remain action-centric and ca

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

Prompt Engineering Strategies for LLM-based Qualitative Coding of Psychological Safety in Software Engineering Communities: A Controlled Empirical Study

DGX agent

arXiv:2605.07422v1 Announce Type: cross Abstract: Qualitative analysis plays a pivotal role in understanding the human and social aspects of software engineering. However, it remains a demanding proce

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

PSK@EEUCA 2026: Fine-Tuning Large Language Models with Synthetic Data Augmentation for Multi-Class Toxicity Detection in Gaming Chat

DGX agent

arXiv:2605.07201v1 Announce Type: cross Abstract: This paper describes our system for the EEUCA 2026 Shared Task on Understanding Toxic Behavior in Gaming Communities. The task involves classifying Wo

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Quality-Conditioned Agreement in Automated Short Answer Scoring: Mid-Range Degradation and the Impact of Task-Specific Adaptation

DGX agent

arXiv:2605.07647v1 Announce Type: cross Abstract: Automated short answer scoring (ASAS) is shifting from discriminative, fine-tuned models to large language models (LLMs) used in few-shot settings. Th

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Query-efficient model evaluation using cached responses

DGX agent

arXiv:2605.07096v1 Announce Type: cross Abstract: Evaluating a new model on an existing benchmark is often necessary to understand its behavior before deployment. For modern evaluation frameworks, gen

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Quotient Semivalues for False-Name-Resistant Data Attribution

DGX agent

arXiv:2605.07663v1 Announce Type: cross Abstract: Data valuation methods allocate payments and audit training data's contribution to machine-learning pipelines; however, they often assume passive cont

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

Qwen3-VL-Seg: Unlocking Open-World Referring Segmentation with Vision-Language Grounding

DGX agent

arXiv:2605.07141v1 Announce Type: cross Abstract: Open-world referring segmentation requires grounding unconstrained language expressions to precise pixel-level regions. Existing multimodal large lang

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Randomness is sometimes necessary for coordination

DGX agent

arXiv:2605.06825v1 Announce Type: new Abstract: Full parameter sharing is standard in cooperative multi-agent reinforcement learning (MARL) for homogeneous agents. Under permutation-symmetric observat

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

RCoT-Seg: Reinforced Chain-of-Thought for Video Reasoning and Segmentation

DGX agent

arXiv:2605.07334v1 Announce Type: new Abstract: Video Reasoning Segmentation (VRS) aims to segment target objects in videos based on implicit instructions that convey human intent and temporal logic.

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

Real-IAD MVN: A Multi-View Normal Vector Dataset and Benchmark for High-Fidelity Industrial Anomaly Detection

DGX agent

arXiv:2605.07149v1 Announce Type: new Abstract: Industrial Anomaly Detection (IAD) is critical for quality control, but existing methods struggle with subtle, geometric defects. Standard 2D (RGB) imag

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

ReasonSTL: Bridging Natural Language and Signal Temporal Logic via Tool-Augmented Process-Rewarded Learning

DGX agent

arXiv:2605.06483v2 Announce Type: replace Abstract: Signal Temporal Logic (STL) is an expressive formal language for specifying spatio-temporal requirements over real-valued, real-time signals. It has

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

RedDiffuser: Auditing Multimodal Safety Failures in Vision-Language Models via Reinforced Diffusion

DGX agent

arXiv:2503.06223v5 Announce Type: replace Abstract: Large Vision-Language Models (VLMs) are increasingly deployed in open-ended environments, where ensuring reliable safety under multimodal inputs is

model-releasesarxiv-cs-cv
11 May 2026
← Previous
1…266267268269270…361
Next →