AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,288 results
Model Releases

PSK at SemEval-2026 Task 9: Multilingual Polarization Detection Using Ensemble Gemma Models with Synthetic Data Augmentation

DGX agent

arXiv:2605.05159v1 Announce Type: new Abstract: We present our system for SemEval-2026 Task 9: Multilingual Polarization Detection, a binary classification task spanning 22 languages. Our approach fin

model-releasesarxiv-cs-cl
7 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

QKVShare: Quantized KV-Cache Handoff for Multi-Agent On-Device LLMs

DGX agent

arXiv:2605.03884v1 Announce Type: new Abstract: Multi-agent LLM systems on edge devices need to hand off latent context efficiently, but the practical choices today are expensive re-prefill or full-pr

model-releasesarxiv-cs-ai
7 May 2026
Model Releases

Quantum-inspired Reinforcement Learning for Synthesizable Drug Design

DGX agent

arXiv:2409.09183v2 Announce Type: replace Abstract: Synthesizable molecular design (also known as synthesizable molecular optimization) is a fundamental problem in drug discovery, and involves designi

model-releasesarxiv-cs-lg
7 May 2026
Model Releases

Real-Time Evaluation of Autonomous Systems under Adversarial Attacks

DGX agent

arXiv:2605.03491v1 Announce Type: new Abstract: Most evaluations of autonomous driving policies under adversarial conditions are conducted in simulation, due to cost efficiency and the absence of phys

model-releasesarxiv-cs-ai
7 May 2026
Model Releases

ReasonAudio: A Benchmark for Evaluating Reasoning Beyond Matching in Text-Audio Retrieval

DGX agent

arXiv:2605.03361v2 Announce Type: new Abstract: As multimodal content continues to expand at a rapid pace, audio retrieval has emerged as a key enabling technology for media search, content organizati

model-releasesarxiv-cs-ai
7 May 2026
Model Releases

Redefining AI Red Teaming in the Agentic Era: From Weeks to Hours

DGX agent

arXiv:2605.04019v1 Announce Type: new Abstract: AI systems are entering critical domains like healthcare, finance, and defense, yet remain vulnerable to adversarial attacks. While AI red teaming is a

model-releasesarxiv-cs-ai
7 May 2026
Model Releases

Regime-Conditioned Evaluation in Multi-Context Bayesian Optimization

DGX agent

arXiv:2605.04895v1 Announce Type: new Abstract: Published transfer-BO comparisons often estimate an average treatment effect of acquisition choice over hidden regime variables, while practitioners nee

model-releasesarxiv-cs-lg
7 May 2026
Model Releases

Replacing Parameters with Preferences: Federated Alignment of Heterogeneous Vision-Language Models

DGX agent

arXiv:2605.03426v1 Announce Type: new Abstract: Vision-Language Models (VLMs) have broad potential in privacy-sensitive domains such as healthcare and finance, yet strict data-sharing constraints rend

model-releasesarxiv-cs-ai
7 May 2026
Model Releases

RLearner-LLM: Balancing Logical Grounding and Fluency in Large Language Models via Hybrid Direct Preference Optimization

DGX agent

arXiv:2605.04539v1 Announce Type: new Abstract: Direct Preference Optimization (DPO), the efficient alternative to PPO-based RLHF, falls short on knowledge-intensive generation: standard preference si

model-releasesarxiv-cs-cl
7 May 2026
Model Releases

RoDyGS: Robust Dynamic Gaussian Splatting for Casual Videos

DGX agent

arXiv:2412.03077v2 Announce Type: replace Abstract: 4D reconstruction from casually captured monocular videos is challenging due to inherent ambiguity in reconstructing dynamic 3D geometry. To address

model-releasesarxiv-cs-cv
7 May 2026
Model Releases

Scalable Object Detection in the Car Interior With Vision Foundation Models

DGX agent

arXiv:2508.19651v2 Announce Type: replace Abstract: AI tasks in the car interior like identifying and localizing externally introduced objects is crucial for response quality of personal assistants. H

model-releasesarxiv-cs-cv
7 May 2026
Model Releases

Self-Attention as Transport: Limits of Symmetric Spectral Diagnostics

DGX agent

arXiv:2605.04893v1 Announce Type: cross Abstract: Large language models hallucinate in predictable ways: attention routing fails by over-concentrating on a narrow set of positions, or by spreading so

model-releasesarxiv-cs-cl
7 May 2026
Model Releases

Self-Prompting Small Language Models for Privacy-Sensitive Clinical Information Extraction

DGX agent

arXiv:2605.04221v1 Announce Type: new Abstract: Clinical named entity recognition from dental progress notes is challenging because documentation is highly unstructured, domain-specific, and often pri

model-releasesarxiv-cs-cl
7 May 2026
Model Releases

Sharp Capacity Thresholds in Linear Associative Memory: From Winner-Take-All to Listwise Retrieval

DGX agent

arXiv:2605.05189v1 Announce Type: cross Abstract: How many key-value associations can a dimes d linear memory store? We show that the answer depends not only on the d^2 degrees of freedom in the memor

model-releasesarxiv-cs-lg
7 May 2026
Model Releases

Single-Position Intervention Fails: Distributed Output Templates Drive In-Context Learning

DGX agent

arXiv:2605.04061v1 Announce Type: cross Abstract: Understanding how large language models encode task identity from few-shot demonstrations is a central open problem in mechanistic interpretability. P

model-releasesarxiv-cs-cl
7 May 2026
Model Releases

SkCC: Portable and Secure Skill Compilation for Cross-Framework LLM Agents

DGX agent

arXiv:2605.03353v1 Announce Type: cross Abstract: LLM-Agents have evolved into autonomous systems for complex task execution, with the SKILL.md specification emerging as a de facto standard for encaps

model-releasesarxiv-cs-ai
7 May 2026
Model Releases

Skill Neologisms: Towards Skill-based Continual Learning

DGX agent

arXiv:2605.04970v1 Announce Type: new Abstract: Modern LLMs show mastery over an ever-growing range of skills, as well as the ability to compose them flexibly. However, extending model capabilities to

model-releasesarxiv-cs-lg
7 May 2026
Model Releases

SlotVLA: Towards Modeling of Object-Relation Representations in Robotic Manipulation

DGX agent

arXiv:2511.06754v3 Announce Type: replace-cross Abstract: Inspired by how humans reason over discrete objects and their relationships, we explore whether compact object-centric and object-relation rep

model-releasesarxiv-cs-cv
7 May 2026
Model Releases

Sparse Autoencoder Decomposition of Clinical Sequence Model Representations: Feature Complexity, Task Specialisation, and Mortality Prediction

DGX agent

arXiv:2605.04072v1 Announce Type: cross Abstract: Sparse autoencoders (SAEs) have been applied to large language models and protein language models, but not systematically to electronic health record

model-releasesarxiv-cs-cl
7 May 2026
Model Releases

SpecPL: Disentangling Spectral Granularity for Prompt Learning

DGX agent

arXiv:2605.04504v1 Announce Type: cross Abstract: Existing prompt learning for VLMs exhibits a modality asymmetry, predominantly optimizing text tokens while still relying on frozen visual encoder as

model-releasesarxiv-cs-cl
7 May 2026
Model Releases

Stable Agentic Control: Tool-Mediated LLM Architecture for Autonomous Cyber Defense

DGX agent

arXiv:2605.03034v1 Announce Type: new Abstract: Agentic systems involved in high-stake decision-making under adversarial pressure need formal guarantees not offered by existing approaches. Motivated b

model-releasesarxiv-cs-ai
7 May 2026
Model Releases

StableI2I: Spotting Unintended Changes in Image-to-Image Transition

DGX agent

arXiv:2605.04453v1 Announce Type: new Abstract: In most real-world image-to-image (I2I) scenarios, existing evaluations primarily focus on instruction following and the perceptual quality or aesthetic

model-releasesarxiv-cs-cv
7 May 2026
Model Releases

Stage Light is Sequence^2: Multi-Light Control via Imitation Learning

DGX agent

arXiv:2605.03660v1 Announce Type: cross Abstract: Music-inspired Automatic Stage Lighting Control (ASLC) has gained increasing attention in recent years due to the substantial time and financial costs

model-releasesarxiv-cs-ai
7 May 2026
Model Releases

Storage Is Not Memory: A Retrieval-Centered Architecture for Agent Recall

DGX agent

arXiv:2605.04897v1 Announce Type: new Abstract: Extraction at ingestion is the wrong primitive for agent memory: content discarded before the query is known cannot be recovered at retrieval time. We p

model-releasesarxiv-cs-cl
7 May 2026
Model Releases

StoryAlign: Evaluating and Training Reward Models for Story Generation

DGX agent

arXiv:2605.04831v1 Announce Type: new Abstract: Story generation aims to automatically produce coherent, structured, and engaging narratives. Although large language models (LLMs) have significantly a

model-releasesarxiv-cs-cl
7 May 2026
Model Releases

SWAN: Semantic Watermarking with Abstract Meaning Representation

DGX agent

arXiv:2605.04305v1 Announce Type: new Abstract: We introduce SWAN (Semantic Watermarking with Abstract Meaning Representation), a novel framework that embeds watermark signatures into the semantic str

model-releasesarxiv-cs-cl
7 May 2026
Model Releases

Syntax- and Compilation-Preserving Evasion of LLM Vulnerability Detectors

DGX agent

arXiv:2602.00305v2 Announce Type: replace-cross Abstract: LLM-based vulnerability detectors are increasingly deployed in CI/CD security gating, yet their resilience to evasion under syntax- and compil

model-releasesarxiv-cs-lg
7 May 2026
Model Releases

TabEmbed: Benchmarking and Learning Generalist Embeddings for Tabular Understanding

DGX agent

arXiv:2605.04962v1 Announce Type: new Abstract: Foundation models have established unified representations for natural language processing, yet this paradigm remains largely unexplored for tabular dat

model-releasesarxiv-cs-cl
7 May 2026
Model Releases

TCM-Serve: Modality-aware Scheduling for Multimodal Large Language Model Inference

DGX agent

arXiv:2603.26498v2 Announce Type: replace-cross Abstract: Multimodal Large Language Models (MLLMs) power platforms like ChatGPT, Gemini, and Copilot, enabling richer interactions with text, images, an

model-releasesarxiv-cs-ai
7 May 2026
Model Releases

Telegraph English: Semantic Prompt Compression via Structured Symbolic Rewriting

DGX agent

arXiv:2605.04426v1 Announce Type: new Abstract: We introduce Telegraph English (TE), a prompt-compression protocol that rewrites natural language into a symbol-rich, formally-structured dialect. Where

model-releasesarxiv-cs-cl
7 May 2026
Model Releases

Terminus-4B: Can a Smaller Model Replace Frontier LLMs at Agentic Execution Tasks?

DGX agent

arXiv:2605.03195v1 Announce Type: new Abstract: Modern coding agents increasingly delegate specialized subtasks to subagents, which are smaller, focused agentic loops that handle narrow responsibiliti

model-releasesarxiv-cs-ai
7 May 2026
Model Releases

The Predictive-Causal Gap: An Impossibility Theorem and Large-Scale Neural Evidence

DGX agent

arXiv:2605.05029v1 Announce Type: new Abstract: We report a systematic failure mode in predictive representation learning. Across 2695 neural network configurations trained to predict linear-Gaussian

model-releasesarxiv-cs-lg
7 May 2026
Model Releases

The Shape of Beliefs: Geometry, Dynamics, and Interventions along Representation Manifolds of Language Models' Posteriors

DGX agent

arXiv:2602.02315v2 Announce Type: replace Abstract: Large language models (LLMs) form implicit beliefs (posteriors over latent variables) from prompts, but we lack a mechanistic account of how these b

model-releasesarxiv-cs-cl
7 May 2026
Model Releases

Tightly-Coupled Estimation and Guidance for Robust Low-Thrust Rendezvous via Adaptive Homotopy

DGX agent

arXiv:2605.04481v1 Announce Type: new Abstract: Minimum-fuel low-thrust rendezvous guidance yields bang-bang control structures highly sensitive to estimation errors, sensor anomalies, and solver regu

model-releasesarxiv-cs-ro
7 May 2026
Model Releases

Transformation Categorization Based on Group Decomposition Theory Using Parameter Division

DGX agent

arXiv:2605.04056v1 Announce Type: new Abstract: Representation learning seeks meaningful sensory representations without supervision and can model aspects of human development. Although many neural ne

model-releasesarxiv-cs-lg
7 May 2026
Model Releases

Tree-Conditioned Edit Flows for Ancestral Sequence Reconstruction

DGX agent

arXiv:2605.04119v1 Announce Type: cross Abstract: Ancestral sequence reconstruction (ASR) aims to infer extinct protein sequences at internal nodes of a phylogenetic tree. Classical ASR methods are ty

model-releasesarxiv-cs-lg
7 May 2026
Model Releases

TSCG: Deterministic Tool-Schema Compilation for Agentic LLM Deployments

DGX agent

arXiv:2605.04107v1 Announce Type: cross Abstract: Production agent frameworks (OpenAI Function Calling, Anthropic Tool Use, MCP) transmit tool schemas as JSON, a format designed for machine parsing, n

model-releasesarxiv-cs-cl
7 May 2026
Model Releases

UAV as Urban Construction Change Monitor: A New Benchmark and Change Captioning Model

DGX agent

arXiv:2605.04409v1 Announce Type: new Abstract: Remote Sensing Image Change Captioning (RSICC) aims to generate spatially grounded natural language descriptions of scene evolution from bi-temporal ima

model-releasesarxiv-cs-cv
7 May 2026
Model Releases

UFAL-CUNI at SemEval-2026 Task 11: An Efficient Modular Neuro-symbolic Method for Syllogistic Reasoning

DGX agent

arXiv:2605.04941v1 Announce Type: new Abstract: This paper describes our system submitted to SemEval-2026 Task 11: Disentangling Content and Formal Reasoning in Large Language Models. We present an ef

model-releasesarxiv-cs-cl
7 May 2026
Model Releases

Unified Framework of Distributional Regret in Multi-Armed Bandits and Reinforcement Learning

DGX agent

arXiv:2605.05102v1 Announce Type: new Abstract: We study the distribution of regret in stochastic multi-armed bandits and episodic reinforcement learning through a unified framework. We formalize a di

model-releasesarxiv-cs-lg
7 May 2026
Model Releases

VCBench: Benchmarking LLMs in Venture Capital

DGX agent

arXiv:2509.14448v2 Announce Type: replace Abstract: Benchmarks such as SWE-bench and ARC-AGI demonstrate how shared datasets accelerate progress toward artificial general intelligence (AGI). We introd

model-releasesarxiv-cs-ai
7 May 2026
Model Releases

Vision-EKIPL: External Knowledge-Infused Policy Learning for Visual Reasoning

DGX agent

arXiv:2506.06856v3 Announce Type: replace Abstract: Visual reasoning is crucial for understanding complex multimodal data and advancing Artificial General Intelligence. Existing methods enhance the re

model-releasesarxiv-cs-cv
7 May 2026
Model Releases

VL-UniTrack: A Unified Framework with Visual-Language Prompts for UAV-Ground Visual Tracking

DGX agent

arXiv:2605.04574v1 Announce Type: new Abstract: UAV-ground visual tracking (UGVT) aims to simultaneously track the same object from both the UAV and the ground view. However, existing two-stream metho

model-releasesarxiv-cs-cv
7 May 2026
Model Releases

VTAgent: Agentic Keyframe Anchoring for Evidence-Aware Video TextVQA

DGX agent

arXiv:2605.04870v1 Announce Type: new Abstract: Video text-based visual question answering (Video TextVQA) aims to answer questions by reasoning over visual textual content appearing in videos. Despit

model-releasesarxiv-cs-cv
7 May 2026
Model Releases

Wasserstein-Aligned Localisation for VLM-Based Distributional OOD Detection in Medical Imaging

DGX agent

arXiv:2605.05161v1 Announce Type: new Abstract: Zero-shot anomaly localisation via vision-language models (VLMs) offers a compelling approach for rare pathology detection, yet its performance is funda

model-releasesarxiv-cs-cv
7 May 2026
Model Releases

What Happens Inside Agent Memory? Circuit Analysis from Emergence to Diagnosis

DGX agent

arXiv:2605.03354v1 Announce Type: new Abstract: Agent memory failures are silent: an LLM-based agent can produce a fluent response even when it fails to extract, retain, or retrieve the information ne

model-releasesarxiv-cs-ai
7 May 2026
Model Releases

12 Angry AI Agents: Evaluating Multi-Agent LLM Decision-Making Through Cinematic Jury Deliberation

DGX agent

arXiv:2605.01986v1 Announce Type: new Abstract: What if the twelve jurors of Sidney Lumet's 12 Angry Men (1957) were not men, but large language models? Would the one juror who disagrees still be able

model-releasesarxiv-cs-ai
6 May 2026
Model Releases

6G Needs Agents: Toward Agentic AI-Native Networks for Autonomous Intelligence

DGX agent

arXiv:2605.01546v1 Announce Type: cross Abstract: Sixth-generation (6G) networks are increasingly envisioned as AI-native infrastructures integrating communication, sensing, and computing into a unifi

model-releasesarxiv-cs-ai
6 May 2026
← Previous
1…271272273274275…361
Next →