AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
84,661Total entries
1Added by human
84,660Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,628 results
Model Releases

DDX-TRACE: A Benchmark for Medical Diagnostic Trajectories in VLMs

DGX agent

arXiv:2605.23629v1 Announce Type: new Abstract: Medical diagnosis is not a single prediction from a fully specified vignette. It is a sequential workup: clinicians decide what evidence to obtain, revi

model-releasesarxiv-cs-cv
25 May 2026
Model Releases
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Decomposing and Measuring Evaluation Awareness

DGX agent

arXiv:2605.23055v1 Announce Type: cross Abstract: Frontier language models sometimes recognize that they are being evaluated and adjust their behavior, undermining validity of benchmark results. Yet t

model-releasesarxiv-cs-ai
25 May 2026
Model Releases

Decomposing Queries into Tool Calls for Long-Video Keyframe Retrieval

DGX agent

arXiv:2605.23826v1 Announce Type: cross Abstract: Keyframe selection is a direct way to provide verifiable visual evidence for long-video question answering (QA). Queries differ in what they require,

model-releasesarxiv-cs-cl
25 May 2026
Model Releases

Decomposition-Based Modular Conformal Prediction for Two-Stage Modeling

DGX agent

arXiv:2510.04406v2 Announce Type: replace-cross Abstract: Conformal prediction offers finite-sample coverage guarantees under minimal assumptions. However, existing methods treat the entire modeling p

model-releasesarxiv-cs-lg
25 May 2026
Model Releases

Decoupling Spatio-Temporal Adapter for Fine-Grained Badminton Action Localization

DGX agent

arXiv:2605.23355v1 Announce Type: new Abstract: Temporal Action Localization (TAL) has been extensively studied in generic video understanding, while fine-grained sports scenarios, such as professiona

model-releasesarxiv-cs-cv
25 May 2026
Model Releases

DeepSeek V4 Flash IS BACK on Nous Portal for FREE for use in Hermes Agent! Check it out at https://portal.nousresearch.com/manage-subscripti…

DGX agent

DeepSeek V4 Flash model has been made available again on the Nous Research portal at no cost for use with Hermes Agent applications. Users can access and utilize this model through the Nous portal's s

model-releasesnous-research--x
25 May 2026
Model Releases

DepthAgent: Towards Better Universal Depth Estimation via Sample-wise Expert Selection

DGX agent

arXiv:2605.23281v1 Announce Type: new Abstract: Monocular metric depth estimation has achieved strong progress with large-scale training and universal-camera modeling, yet robust deployment across div

model-releasesarxiv-cs-cv
25 May 2026
Model Releases

Design and Report Benchmarks for Knowledge Work

DGX agent

arXiv:2605.23262v1 Announce Type: new Abstract: The development of LLM agents has led to a growing body of work on knowledge-work AI, including coding, research, and healthcare. However, current knowl

model-releasesarxiv-cs-ai
25 May 2026
Model Releases

DFKI-MLT at SemEval-2026 TASK 7: Steering Multilingual Models Towards Cultural Knowledge

DGX agent

arXiv:2605.23069v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used across diverse linguistic and cultural contexts, yet their cultural knowledge remains uneven across r

model-releasesarxiv-cs-cl
25 May 2026
Model Releases

Diffusion-based Denoising Beats Vanilla Score Matching in Parameter Estimation: A Theoretical Explanation

DGX agent

arXiv:2605.22950v1 Announce Type: cross Abstract: Score matching is an alternative to maximum likelihood estimation when the normalizing constant is unknown or too costly to evaluate. However, vanilla

model-releasesarxiv-cs-lg
25 May 2026
Model Releases

Discontinuous Galerkin Neural Operator for Pathology Defocus Deblurring

DGX agent

arXiv:2605.23282v1 Announce Type: cross Abstract: Defocus deblurring in pathological microscopy remains challenging due to the spatially varying and locally discontinuous nature of optical blur induce

model-releasesarxiv-cs-cv
25 May 2026
Model Releases

Do Synthetic Brain MRIs Reliably Improve Tumour Classification? A StyleGAN2-ADA Class-Plane Augmentation Study on BRISC 2025

DGX agent

arXiv:2605.23094v1 Announce Type: cross Abstract: Generative augmentation is often proposed as a remedy for small medical-image datasets, but synthetic images are only useful when they improve downstr

model-releasesarxiv-cs-ai
25 May 2026
Model Releases

DreamerNLplus: Interpretable Modeling of Mental Health Dynamics from Social Media Timelines using Hybrid Rule-Based and RAG Methods

DGX agent

arXiv:2605.23052v1 Announce Type: cross Abstract: We present DreamerNLplus, a hybrid framework for modeling mental health dynamics from social media timelines in the CLPsych 2026 shared task. Our syst

model-releasesarxiv-cs-ai
25 May 2026
Model Releases

DRIVESPATIAL: A Benchmark for Spatiotemporal Intelligence in VLMs for Autonomous Driving

DGX agent

arXiv:2605.23176v1 Announce Type: new Abstract: Spatiotemporal intelligence in autonomous driving (AD) requires an agent to integrate multi-view observations into a coherent scene representation, main

model-releasesarxiv-cs-cv
25 May 2026
Model Releases

Efficient and Transferable Agentic Knowledge Graph RAG via Reinforcement Learning

DGX agent

arXiv:2509.26383v5 Announce Type: replace-cross Abstract: Knowledge-graph retrieval-augmented generation (KG-RAG) couples large language models (LLMs) with structured, verifiable knowledge graphs (KGs

model-releasesarxiv-cs-ai
25 May 2026
Model Releases

Efficient Gradient Estimation for Parameterized Quantum Systems with Lie Algebraic Symmetries

DGX agent

arXiv:2404.05108v3 Announce Type: replace-cross Abstract: Gradient estimation is a central challenge in training parameterized quantum circuits (PQCs) for hybrid quantum-classical optimization and lea

model-releasesarxiv-cs-lg
25 May 2026
Model Releases

Efficient One-Step Diffusion Restoration Model with Compact Token Compression and Linear Attention

DGX agent

arXiv:2605.23451v1 Announce Type: new Abstract: Real-world image super-resolution aims to recover high-quality images from complex and unknown real-world degradations. However, existing generative Rea

model-releasesarxiv-cs-cv
25 May 2026
Model Releases

Entropy Equivalence Testing

DGX agent

arXiv:2605.23225v1 Announce Type: cross Abstract: We introduce the problem of entropy equivalence testing for probability distributions, a relaxation of the well-studied closeness testing problem, whe

model-releasesarxiv-cs-lg
25 May 2026
Model Releases

ETCHR: Editing To Clarify and Harness Reasoning

DGX agent

arXiv:2605.23897v1 Announce Type: cross Abstract: Multimodal Large Language Models have advanced visual reasoning, yet a purely textual chain of thought remains a bottleneck for questions that require

model-releasesarxiv-cs-ai
25 May 2026
Model Releases

Evaluating Large Language Models in a Complex Hidden Role Game

DGX agent

arXiv:2605.22826v1 Announce Type: cross Abstract: Quantifying the deceptive potential of Large Language Models (LLMs) is critical for AI safety, yet difficult to achieve in uncontrolled environments.

model-releasesarxiv-cs-ai
25 May 2026
Model Releases

Evaluating Memory Structure in LLM Agents

DGX agent

arXiv:2602.11243v2 Announce Type: replace-cross Abstract: Modern LLM-based agents and chat assistants rely on long-term memory frameworks to store reusable knowledge, recall user preferences, and augm

model-releasesarxiv-cs-cl
25 May 2026
Model Releases

Exploitation of KnowledgeDeliver via ViewState Deserialization Vulnerability

DGX agent

Written by: Takahiro Sugiyama, Peter Revelant, Mathew Potaczek Introduction In late 2025, Mandiant responded to a security incident involving a compromised web server running KnowledgeDeliver. Knowled

model-releasesgoogle-cloud-ai
25 May 2026
Model Releases

Exploiting Longitudinal Context in Clinician-Verified Interactive Lesion Tracking

DGX agent

arXiv:2605.23118v1 Announce Type: cross Abstract: Tracking tumor lesions across serial CT scans is essential for oncological response assessment. Existing automated methods face a fundamental trade-of

model-releasesarxiv-cs-ai
25 May 2026
Model Releases

Exploring deep learning for Event-Based Saliency Prediction with a Transformer-based model

DGX agent

arXiv:2605.23790v1 Announce Type: new Abstract: Saliency prediction has been extensively studied in RGB images and videos as a computational model of human visual attention. In contrast, predicting sa

model-releasesarxiv-cs-cv
25 May 2026
Model Releases

FAST-ME: Foundation-aware Adaptive Stopping for Motion Estimation for Efficient IoT Video Analysis

DGX agent

arXiv:2605.23428v1 Announce Type: new Abstract: In modern multimedia systems, efficient video processing is critical, especially in resource-constrained environments such as IoT-based camera networks,

model-releasesarxiv-cs-cv
25 May 2026
Model Releases

FastKernels: Benchmarking GPU Kernel Generation in Production

DGX agent

arXiv:2605.23215v1 Announce Type: cross Abstract: LLM-based agents for GPU kernel generation are advancing rapidly, yet their progress is fundamentally constrained by the benchmarks they optimize agai

model-releasesarxiv-cs-ai
25 May 2026
Model Releases

FATHOMS-RAG: A Framework for the Assessment of Thinking and Observation in Multimodal Systems that use Retrieval Augmented Generation

DGX agent

arXiv:2510.08945v3 Announce Type: replace Abstract: Retrieval-augmented generation (RAG) has emerged as a promising paradigm for improving factual accuracy in large language models (LLMs). We introduc

model-releasesarxiv-cs-ai
25 May 2026
Model Releases

Fine-Tuning Causal LLMs for Text Classification: Embedding-Based vs. Instruction-Based Approaches

DGX agent

arXiv:2512.12677v2 Announce Type: replace-cross Abstract: We explore efficient strategies to fine-tune decoder-only Large Language Models (LLMs) for downstream text classification under resource const

model-releasesarxiv-cs-ai
25 May 2026
Model Releases

FuRA: Full-Rank Parameter-Efficient Fine-Tuning with Spectral Preconditioning

DGX agent

arXiv:2605.22869v1 Announce Type: new Abstract: Both full fine-tuning (Full FT) and parameter-efficient fine-tuning methods such as LoRA introduce weight updates without accounting for the spectral st

model-releasesarxiv-cs-lg
25 May 2026
Model Releases

GENSTRAT: Toward a Science of Strategic Reasoning in Large Language Models

DGX agent

arXiv:2605.23238v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed as economic agents in marketplaces, auctions, and bidding settings. Anticipating their behavior i

model-releasesarxiv-cs-ai
25 May 2026
Model Releases

Given how much of the original 'bottle of water per generated email' water estimate came from guesses at the architecture of GPT-4, it would…

DGX agent

Given how much of the original 'bottle of water per generated email' water estimate came from guesses at the architecture of GPT-4, it would be very much in @OpenAI's interest to publish the architect

model-releasessimon-willison--x
25 May 2026
Model Releases

. @GrooveJonesXR needed to deliver the impossible: giant NFL-licensed Crocs parachuting into Dick's Sporting Goods parking lots; hyper-reali…

DGX agent

. @GrooveJonesXR needed to deliver the impossible: giant NFL-licensed Crocs parachuting into Dick's Sporting Goods parking lots; hyper-realistic, multi-location, vertical 9:16 on a holiday deadline. A

model-releasescomfyui--x
25 May 2026
Model Releases

GT-HarmBench: Benchmarking AI Safety Risks Through the Lens of Game Theory

DGX agent

arXiv:2602.12316v2 Announce Type: replace Abstract: Frontier AI systems are increasingly capable and deployed in high-stakes multi-agent environments. However, existing AI safety benchmarks largely ev

model-releasesarxiv-cs-ai
25 May 2026
Model Releases

HARNESS-LM: A Three-Phase Training Recipe for Harnessing SLMs in Sponsored Search Retrieval

DGX agent

arXiv:2605.23572v1 Announce Type: cross Abstract: In the competitive landscape of sponsored search, balancing retrieval quality with production latency is a critical challenge. While large retrieval m

model-releasesarxiv-cs-ai
25 May 2026
Model Releases

Hermes Agent now can orchestrate the @OpenHandsDev agents with a new optional skill! `hermes update` then do `hermes skills install official…

DGX agent

Hermes Agent now can orchestrate the @OpenHandsDev agents with a new optional skill! `hermes update` then do `hermes skills install official/autonomous-ai-agents/openhands` Reminder: You can already d

model-releasesnous-research--x
25 May 2026
Model Releases

Hierarchical Concept Geometry in Language Models Emerges from Word Co-occurrence

DGX agent

arXiv:2605.23821v1 Announce Type: new Abstract: We propose a distributional theory of how hypernymy -- the ``is-a'' relation between general and specific concepts -- is encoded geometrically in langua

model-releasesarxiv-cs-cl
25 May 2026
Model Releases

How Far Are We from Generating Missing Modalities with Foundation Models?

DGX agent

arXiv:2506.03530v3 Announce Type: replace-cross Abstract: Multimodal foundation models have demonstrated impressive capabilities across diverse tasks. However, their potential as plug-and-play solutio

model-releasesarxiv-cs-cl
25 May 2026
Model Releases

How Hard is it to Rig a Benchmark? A Social Choice Analysis of Leaderboard Robustness

DGX agent

arXiv:2605.23628v1 Announce Type: new Abstract: Multi-task benchmarks have become a central pillar of machine learning research, yet their growing influence has incentivised benchmark gaming -- strate

model-releasesarxiv-cs-lg
25 May 2026
Model Releases

HTMuon: Improving Muon via Heavy-Tailed Spectral Correction

DGX agent

arXiv:2603.10067v2 Announce Type: replace-cross Abstract: Muon has recently shown promising results in LLM training. In this work, we study how to further improve Muon. We argue that Muon's orthogonal

model-releasesarxiv-cs-ai
25 May 2026
Model Releases

Human-Centered Learning Mechanics: A Dynamical Framework for Entropy-Regulated Representation Learning

DGX agent

arXiv:2605.22940v1 Announce Type: cross Abstract: Deep learning is increasingly viewed as a dynamical process in parameter space, yet many existing theories still treat training as a closed optimizati

model-releasesarxiv-cs-ai
25 May 2026
Model Releases

✅Implicit caching is now live on Qwen3.7-Max — kicks in automatically, no setup needed. ⚡️Faster + cheaper out of the box. Need higher, more…

DGX agent

✅Implicit caching is now live on Qwen3.7-Max — kicks in automatically, no setup needed. ⚡️Faster + cheaper out of the box. Need higher, more deterministic hit rates? Try explicit caching instead. 🙌 🔗B

model-releasesqwen--x
25 May 2026
Model Releases

ImProver 2: Iteratively Self-Improving LMs for Neurosymbolic Proof Optimization

DGX agent

arXiv:2605.22885v1 Announce Type: new Abstract: Formal mathematics libraries are rapidly expanding, creating a growing need to refactor verified proofs for maintainability and to improve training data

model-releasesarxiv-cs-ai
25 May 2026
Model Releases

In his ~43,000-word encyclical, the Pope urged governments to slow down AI development and decried 'new forms of slavery' in AI and tech supply chains (Joshua McElwee/Reuters)

DGX agent

Joshua McElwee / Reuters: In his ~43,000-word encyclical, the Pope urged governments to slow down AI development and decried “new forms of slavery” in AI and tech supply chains — Pope Leo urged govern

model-releasestechmeme
25 May 2026
Model Releases

Inductive Deductive Synthesis: Enabling AI to Generate Formally Verified Systems

DGX agent

arXiv:2605.23109v1 Announce Type: new Abstract: AI agents increasingly excel at generating, testing, and refining code. However, they fall short on tasks requiring formal guarantees of full coverage t

model-releasesarxiv-cs-ai
25 May 2026
Model Releases

IntentionNav: A Benchmark for Intent-Driven Object Navigation from Implicit Human Instruction

DGX agent

arXiv:2605.23187v1 Announce Type: new Abstract: Existing object navigation benchmarks usually tell an embodied agent which object category to find, such as microwave or chair. Human-facing embodied AI

model-releasesarxiv-cs-cv
25 May 2026
Model Releases

Investigating Robot Control Policy Learning for Autonomous X-ray-guided Spine Procedures

DGX agent

arXiv:2511.03882v2 Announce Type: replace-cross Abstract: Imitation learning-based robot control policies are enjoying renewed interest in video-based robotics. However, it remains unclear whether thi

model-releasesarxiv-cs-ai
25 May 2026
Model Releases

Is Capability a Liability? More Capable Language Models Make Worse Forecasts When It Matters Most

DGX agent

arXiv:2605.22672v2 Announce Type: replace Abstract: We document inverse scaling in LLMs on forecasting problems whose underlying time series exhibit superlinear growth and tail risk of regime change,

model-releasesarxiv-cs-ai
25 May 2026
Model Releases

It's the humans, not the data: Geopolitical bias in LLMs originates in post-training, amplified by the language of the prompt

DGX agent

arXiv:2605.23825v1 Announce Type: cross Abstract: It has generally been assumed that geopolitical bias in language models originates from the training data used during the pre-training phase. We teste

model-releasesarxiv-cs-ai
25 May 2026
← Previous
1…271272273274275…472
Next →