AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,585 results
Model Releases

Prompt-Activation Duality: Improving Activation Steering via Attention-Level Interventions

DGX agent

arXiv:2605.10664v1 Announce Type: cross Abstract: Activation steering controls language model behavior by adding directions to internal representations at inference time, but standard residual-stream

model-releasesarxiv-cs-ai
12 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Prompt Estimation from Prototypes for Federated Prompt Tuning of Vision Transformers

DGX agent

arXiv:2510.25372v2 Announce Type: replace Abstract: Visual Prompt Tuning (VPT) of pre-trained Vision Transformers (ViTs) has proven highly effective as a parameter-efficient fine-tuning technique for

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

PumpSense: Real-Time Detection and Target Extraction of Crypto Pump-and-Dumps on Telegram

DGX agent

arXiv:2605.09431v1 Announce Type: new Abstract: Cryptocurrency pump-and-dump schemes coordinated via Telegram threaten market integrity. However, existing research addressing this specific threat has

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

QM-ToT: A Medical Tree of Thoughts Reasoning Framework for Quantized Model

DGX agent

arXiv:2504.12334v2 Announce Type: replace Abstract: Large language models (LLMs) face significant challenges in specialized biomedical tasks due to the inherent complexity of medical reasoning and the

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

Quantifying Concentration Phenomena of Mean-Field Transformers in the Low-Temperature Regime

DGX agent

arXiv:2605.10931v1 Announce Type: cross Abstract: Transformers with self-attention modules as their core components have become an integral architecture in modern large language and foundation models.

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Quantifying the Utility of User Simulators for Building Collaborative LLM Assistants

DGX agent

arXiv:2605.09808v1 Announce Type: new Abstract: User simulators are increasingly leveraged to build interactive AI assistants, yet how to measure the quality of these simulators remains an open questi

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

Quantitative Sobolev Approximation Bounds for Neural Operators with Empirical Validation on Burgers Equation

DGX agent

arXiv:2605.08170v1 Announce Type: new Abstract: Neural operators have emerged as a powerful tool for learning mappings between infinite-dimensional function spaces. However, their approximation proper

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Quantum Circuit Simulation of Compartmental Drug Dynamics: Leveraging Variational Algorithms for Nonlinear Mixed-Effects Population Pharmacokinetics

DGX agent

arXiv:2605.09691v1 Announce Type: new Abstract: Population pharmacokinetic/pharmacodynamic (PK/PD) modeling traditionally relies on classical ordinary differential equations to simulate drug dynamics.

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Quasi-Linear ICA for Motor Unit Decomposition during Dynamic Contractions

DGX agent

arXiv:2406.19581v2 Announce Type: replace-cross Abstract: Decomposing surface electromyography (EMG) into the spike trains of individual motor neurons is a long-standing inverse problem and a key step

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Queryable LoRA: Instruction-Regularized Routing Over Shared Low-Rank Update Atoms

DGX agent

arXiv:2605.08423v1 Announce Type: cross Abstract: We present a data-adaptive method for parameter-efficient fine-tuning of large neural networks. Standard low-rank adaptation methods improve efficienc

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

Qwen Goes Brrr: Off-the-Shelf RAG for Ukrainian Multi-Domain Document Understanding

DGX agent

arXiv:2605.10296v1 Announce Type: cross Abstract: We participated in the Fifth UNLP shared task on multi-domain document understanding, where systems must answer Ukrainian multiple-choice questions fr

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Qwen-Image-2.0 Technical Report

DGX agent

arXiv:2605.10730v1 Announce Type: new Abstract: We present Qwen-Image-2.0, an omni-capable image generation foundation model that unifies high-fidelity generation and precise image editing within a si

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

R4Det: 4D Radar-Camera Fusion for High-Performance 3D Object Detection

DGX agent

arXiv:2603.11566v2 Announce Type: replace Abstract: 4D radar-camera sensing configuration has gained increasing importance in autonomous driving. However, existing 3D object detection methods that fus

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

RadThinking: A Dataset for Longitudinal Clinical Reasoning in Radiology

DGX agent

arXiv:2605.10761v1 Announce Type: new Abstract: Cancer screening is a reasoning task. A radiologist observes findings, compares them to prior scans, integrates clinical context, and reaches a diagnost

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

RareCP: Regime-Aware Retrieval for Efficient Conformal Prediction

DGX agent

arXiv:2605.08857v1 Announce Type: new Abstract: Recent advances in uncertainty quantification for time series forecasting show that conformal prediction can provide reliable prediction intervals, yet

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

RDEx-CASK: Cauchy Mutation, Archive, and Stagnation Kick for RDEx-CSOP

DGX agent

arXiv:2605.09652v1 Announce Type: cross Abstract: We extend RDEx-CSOP with 3 changes that target stagnation & late-stage variance, plus minor parameter tuning. The second scale factor in the standard

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Re^2Math: Benchmarking Theorem Retrieval in Research-Level Mathematics

DGX agent

arXiv:2605.09012v1 Announce Type: new Abstract: Large language models are increasingly capable at closed-world mathematical reasoning, but research assistance also requires source-grounded use of the

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Real vs. Semi-Simulated: Rethinking Evaluation for Treatment Effect Estimation

DGX agent

arXiv:2605.10430v1 Announce Type: cross Abstract: Estimating heterogeneous treatment effects with machine learning has attracted substantial attention in both academic research and industrial practice

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Really amazing dissection. LLMs hitting rarely. Tools list is amazing.

DGX agent

Really amazing dissection. LLMs hitting rarely. Tools list is amazing. 🤩🤯🤩 Claude Code (still not AGI but biggest advance since GPT-4) is the most neurosymbolic thing I have ever seen in my life. 53 s

model-releasesgary-marcus--x
12 May 2026
Model Releases

ReaMOT: A Benchmark and Framework for Reasoning-based Multi-Object Tracking

DGX agent

arXiv:2505.20381v4 Announce Type: replace Abstract: Referring Multi-Object Tracking (RMOT) aims to track targets specified by language instructions. However, existing RMOT paradigms heavily rely on ex

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

REAP: Automatic Curation of Coding Agent Benchmarks from Interactive Production Usage

DGX agent

arXiv:2604.01527v3 Announce Type: replace-cross Abstract: Production deployment of AI coding agents requires fast, reproducible evaluation signals. Existing industrial practices trade off speed and fi

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Reasoning emerges from constrained inference manifolds in large language models

DGX agent

arXiv:2605.08142v1 Announce Type: cross Abstract: Reasoning in large language models is predominantly evaluated through labeled benchmarks, conflating task performance with the quality of internal inf

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

Recovering Physical Dynamics from Discrete Observations via Intrinsic Differential Consistency

DGX agent

arXiv:2605.08454v1 Announce Type: cross Abstract: Recovering continuous-time dynamics from discrete observations is difficult because local supervision (e.g., pointwise regression targets, derivative

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Recursive Language Models

DGX agent

arXiv:2512.24601v3 Announce Type: replace Abstract: We study allowing large language models (LLMs) to process arbitrarily long prompts through the lens of inference-time scaling. We propose Recursive

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

REI-Bench: Can Embodied Agents Understand Vague Human Instructions in Task Planning?

DGX agent

arXiv:2505.10872v4 Announce Type: replace-cross Abstract: Robot task planning decomposes human instructions into executable action sequences that enable robots to complete a series of complex tasks. A

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Reinforcement Learning Measurement Model

DGX agent

arXiv:2605.09305v1 Announce Type: cross Abstract: Interactive assessments generate sequential process data that are not well handled by conventional item response models. Existing MDP-based measuremen

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Relative Kinetic Utility for Reasoning-Aware Structural Pruning in Large Language Models

DGX agent

arXiv:2605.09008v1 Announce Type: cross Abstract: Chain-of-Thought (CoT) prompting symbolized a huge improvement of reasoning capabilities of Large Language Models (LLMs). However, scaling up test-tim

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

RelBench v2: A Large-Scale Benchmark and Repository for Relational Data

DGX agent

arXiv:2602.12606v2 Announce Type: replace Abstract: Relational deep learning (RDL) has emerged as a powerful paradigm for learning directly on relational databases by modeling entities and their relat

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Reliable LLM-Based Edge-Cloud-Expert Cascades for Telecom Knowledge Systems

DGX agent

arXiv:2512.20012v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are emerging as key enablers of automation in domains such as telecommunications, assisting with tasks including

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Rennala MVR: Improved Time Complexity for Parallel Stochastic Optimization via Momentum-Based Variance Reduction

DGX agent

arXiv:2605.08871v1 Announce Type: cross Abstract: Large-scale machine learning models are trained on clusters of machines that exhibit heterogeneous performance due to hardware variability, network de

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

ReorgGS: Equivalent Distribution Reorganization for 3D Gaussian Splatting

DGX agent

arXiv:2605.08739v1 Announce Type: new Abstract: A converged 3D Gaussian Splatting (3DGS) model may approximate the target scene while remaining poorly parameterized for further optimization. We identi

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

Repeated-Token Counting Reveals a Dissociation Between Representations and Outputs

DGX agent

arXiv:2605.09239v1 Announce Type: new Abstract: Large language models fail at counting repeated tokens despite strong performance on broader reasoning benchmarks. These failures are commonly attribute

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

ReplaySCM: A Benchmark for Executable Causal Mechanism Induction from Interventions

DGX agent

arXiv:2605.08197v1 Announce Type: cross Abstract: Most causal benchmarks for language models score local answers or graph structure. We introduce ReplaySCM, a 1,300 item benchmark for executable causa

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Rethinking Agentic Search with Pi-Serini: Is Lexical Retrieval Sufficient?

DGX agent

arXiv:2605.10848v1 Announce Type: cross Abstract: Does a lexical retriever suffice as large language models (LLMs) become more capable in an agentic loop? This question naturally arises when building

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Rethinking Random Transformers as Adaptive Sequence Smoothers for Sleep Staging

DGX agent

arXiv:2605.09905v1 Announce Type: cross Abstract: Automatic sleep staging commonly adopts Transformers under the assumption that they learn complex long-range dependencies. We challenge this view by r

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Retrieval Mechanisms Surpass Long-Context Scaling in Time Series Forecasting

DGX agent

arXiv:2605.08217v1 Announce Type: new Abstract: Time Series Foundation Models (TSFMs) have borrowed the long context paradigm from natural language processing under the premise that feeding more histo

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Retrieve-then-Steer: Online Success Memory for Test-Time Adaptation of Generative VLAs

DGX agent

arXiv:2605.10094v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models show strong potential for general-purpose robotic manipulation, yet their closed-loop reliability often degrades u

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

RewardHarness: Self-Evolving Agentic Post-Training

DGX agent

arXiv:2605.08703v1 Announce Type: new Abstract: Evaluating instruction-guided image edits requires rewards that reflect subtle human preferences, yet current reward models typically depend on large-sc

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

RoboMemArena: A Comprehensive and Challenging Robotic Memory Benchmark

DGX agent

arXiv:2605.10921v1 Announce Type: new Abstract: Memory is a critical component of robotic intelligence, as robots must rely on past observations and actions to accomplish long-horizon tasks in partial

model-releasesarxiv-cs-ro
12 May 2026
Model Releases

Robust Server Defense Against Unreliable Clients in One-Shot Fair Collaborative Machine Learning

DGX agent

arXiv:2605.08616v1 Announce Type: new Abstract: Collaborative machine learning (CML) enables multiple clients to train a global model jointly in a data-distributed setting. To address data privacy and

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Robust Spectral Watermark for Synthetic Tabular Data

DGX agent

arXiv:2511.21600v2 Announce Type: replace-cross Abstract: The rise of generative AI has enabled the production of high-fidelity synthetic tabular data across fields such as healthcare, finance, and pu

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

ROM: Real-time Overthinking Mitigation via Streaming Detection and Intervention

DGX agent

arXiv:2603.22016v2 Announce Type: replace-cross Abstract: Large Reasoning Models (LRMs) often reach a correct solution before their long Chain-of-Thought trace ends, yet continue with redundant verifi

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Route Before Retrieve: Activating Latent Routing Abilities of LLMs for RAG vs. Long-Context Selection

DGX agent

arXiv:2605.10235v1 Announce Type: new Abstract: Recent advances in large language models (LLMs) have expanded the context window to beyond 128K tokens, enabling long-document understanding and multi-s

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

RubricRefine: Improving Tool-Use Agent Reliability with Training-Free Pre-Execution Refinement

DGX agent

arXiv:2605.09730v1 Announce Type: new Abstract: Iterative self-refinement is a popular inference-time reliability technique, but its effectiveness in code-mode tool use depends heavily on the structur

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

RW-Post: Auditable Evidence-Grounded Multimodal Fact-Checking in the Wild

DGX agent

arXiv:2605.10357v1 Announce Type: cross Abstract: Multimodal misinformation increasingly leverages visual persuasion, where repurposed or manipulated images strengthen misleading text. We introduce ex

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

S2FT: Parameter-Efficient Fine-Tuning in Sparse Spectrum Domain

DGX agent

arXiv:2605.08589v1 Announce Type: new Abstract: Parameter Efficient Fine-Tuning (PEFT) is a key technique for adapting a large pretrained model to downstream tasks by fine-tuning only a small number o

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

SACHI: Structured Agent Coordination via Holistic Information Integration in Multi-Agent Reinforcement Learning

DGX agent

arXiv:2605.08391v1 Announce Type: new Abstract: Cooperative multi-agent reinforcement learning agents that act on partial local observations face a fundamental information bottleneck: the knowledge ne

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

SAFA-SNN: Sparsity-Aware On-Device Few-Shot Class-Incremental Learning with Fast-Adaptive Structure of Spiking Neural Network

DGX agent

arXiv:2510.03648v2 Announce Type: replace Abstract: Continuous learning of novel classes is crucial for edge devices to preserve data privacy and maintain reliable performance in dynamic environments.

model-releasesarxiv-cs-lg
12 May 2026
← Previous
1…334335336337338…471
Next →