AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,280 results
Model Releases

SpatialScore: Towards Comprehensive Evaluation for Spatial Intelligence

DGX agent

arXiv:2505.17012v3 Announce Type: replace-cross Abstract: Existing evaluations of multimodal large language models (MLLMs) on spatial intelligence are typically fragmented and limited in scope. In thi

model-releasesarxiv-cs-ai
14 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

SpecMoE: A Fast and Efficient Mixture-of-Experts Inference via Self-Assisted Speculative Decoding

DGX agent

arXiv:2604.10152v1 Announce Type: new Abstract: The Mixture-of-Experts (MoE) architecture has emerged as a promising approach to mitigate the rising computational costs of large language models (LLMs)

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

SPEED-Bench: A Unified and Diverse Benchmark for Speculative Decoding

DGX agent

arXiv:2604.09557v1 Announce Type: cross Abstract: Speculative Decoding (SD) has emerged as a critical technique for accelerating Large Language Model (LLM) inference. Unlike deterministic system optim

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Spoiler Alert: Narrative Forecasting as a Metric for Tension in LLM Storytelling

DGX agent

arXiv:2604.09854v1 Announce Type: new Abstract: LLMs have so far failed both to generate consistently compelling stories and to recognize this failure--on the leading creative-writing benchmark (EQ-Be

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

SRBench: A Comprehensive Benchmark for Sequential Recommendation with Large Language Models

DGX agent

arXiv:2604.09553v1 Announce Type: cross Abstract: LLM development has aroused great interest in Sequential Recommendation (SR) applications. However, comprehensive evaluation of SR models remains lack

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

StableTTA: Training-Free Test-Time Adaptation that Improves Model Accuracy on ImageNet1K to 96%

DGX agent

arXiv:2604.04552v2 Announce Type: replace-cross Abstract: Ensemble methods are widely used to improve predictive performance, but their effectiveness often comes at the cost of increased memory usage

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

STaR-DRO: Stateful Tsallis Reweighting for Group-Robust Structured Prediction

DGX agent

arXiv:2604.09737v1 Announce Type: cross Abstract: Structured prediction requires models to generate ontology-constrained labels, grounded evidence, and valid structure under ambiguity, label skew, and

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

STARS: Skill-Triggered Audit for Request-Conditioned Invocation Safety in Agent Systems

DGX agent

arXiv:2604.10286v1 Announce Type: new Abstract: Autonomous language-model agents increasingly rely on installable skills and tools to complete user tasks. Static skill auditing can expose capability s

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

StarVLA-alpha: Reducing Complexity in Vision-Language-Action Systems

DGX agent

arXiv:2604.11757v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have recently emerged as a promising paradigm for building general-purpose robotic agents. However, the VLA landsc

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

STORM: End-to-End Referring Multi-Object Tracking in Videos

DGX agent

arXiv:2604.10527v1 Announce Type: cross Abstract: Referring multi-object tracking (RMOT) is a task of associating all the objects in a video that semantically match with given textual queries or refer

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

StyleBench: Evaluating thinking styles in Large Language Models

DGX agent

arXiv:2509.20868v2 Announce Type: replace-cross Abstract: Structured reasoning can improve the inference performance of large language models (LLMs), but it also introduces computational cost and cont

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Sub-32B open weights models now offer GPT-5 level intelligence with Qwen3.5 27B (Reasoning) matching GPT-5 (medium) at 42 and Gemma 4 31B (R…

DGX agent

Sub-32B open weights models now offer GPT-5 level intelligence with Qwen3.5 27B (Reasoning) matching GPT-5 (medium) at 42 and Gemma 4 31B (Reasoning) matching GPT-5 (low) at 39 on the Artificial Analy

model-releasesclem-delangue--x
14 Apr 2026
Model Releases

Switch-JustDance: Benchmarking Whole Body Motion Tracking Controllers Using a Commercial Console Game

DGX agent

arXiv:2511.17925v3 Announce Type: replace-cross Abstract: Recent advances in whole-body robot control have enabled humanoid and legged robots to perform increasingly agile and coordinated motions. How

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

SymphoMotion: Joint Control of Camera Motion and Object Dynamics for Coherent Video Generation

DGX agent

arXiv:2604.03723v2 Announce Type: replace Abstract: Controlling both camera motion and object dynamics is essential for coherent and expressive video generation, yet current methods typically handle o

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

Synthius-Mem: Brain-Inspired Hallucination-Resistant Persona Memory Achieving 94.4% Memory Accuracy and 99.6% Adversarial Robustness on LoCoMo

DGX agent

arXiv:2604.11563v1 Announce Type: cross Abstract: Providing AI agents with reliable long-term memory that does not hallucinate remains an open problem. Current approaches to memory for LLM agents -- s

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

TAG-Head: Time-Aligned Graph Head for Plug-and-Play Fine-grained Action Recognition

DGX agent

arXiv:2604.11498v1 Announce Type: new Abstract: Fine-grained human action recognition (FHAR) is challenging because visually similar actions differ by subtle spatio-temporal cues. Many recent systems

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

Tail-Aware Information-Theoretic Generalization for RLHF and SGLD

DGX agent

arXiv:2604.10727v1 Announce Type: cross Abstract: Classical information-theoretic generalization bounds typically control the generalization gap through KL-based mutual information and therefore rely

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Teaching Language Models How to Code Like Learners: Conversational Serialization for Student Simulation

DGX agent

arXiv:2604.10720v1 Announce Type: new Abstract: Artificial models that simulate how learners act and respond within educational systems are a promising tool for evaluating tutoring strategies and feed

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Template-assisted Contrastive Learning of Task-oriented Dialogue Sentence Embeddings

DGX agent

arXiv:2305.14299v3 Announce Type: replace-cross Abstract: Learning high quality sentence embeddings from dialogues has drawn increasing attentions as it is essential to solve a variety of dialogue-ori

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

TempusBench: An Evaluation Framework for Time-Series Forecasting

DGX agent

arXiv:2604.11529v1 Announce Type: new Abstract: Foundation models have transformed natural language processing and computer vision, and a rapidly growing literature on time-series foundation models (T

model-releasesarxiv-cs-lg
14 Apr 2026
Model Releases

Text-to-Image Models and Their Representation of People from Different Nationalities Engaging in Activities

DGX agent

arXiv:2504.06313v5 Announce Type: replace Abstract: This paper investigates how popular text-to-image (T2I) models, DALL-E 3 and Gemini 3 Pro Preview, depict people from 206 nationalities when prompte

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

The Amazing Agent Race: Strong Tool Users, Weak Navigators

DGX agent

arXiv:2604.10261v1 Announce Type: new Abstract: Existing tool-use benchmarks for LLM agents are overwhelmingly linear: our analysis of six benchmarks shows 55 to 100% of instances are simple chains of

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

The Blind Spot of Agent Safety: How Benign User Instructions Expose Critical Vulnerabilities in Computer-Use Agents

DGX agent

arXiv:2604.10577v1 Announce Type: cross Abstract: Computer-use agents (CUAs) can now autonomously complete complex tasks in real digital environments, but when misled, they can also be used to automat

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

The LLM tunes its own llama.cpp flags (+54% tok/s on Qwen3.5-27B)

DGX agent

This r/ollama post describes a technique where an LLM is used to automatically tune its own llama.cpp runtime flags — such as parameters related to GPU offloading, KV cache quantization, batch sizes,

model-releasesr-ollama
14 Apr 2026
Model Releases

The Missing Knowledge Layer in Cognitive Architectures for AI Agents

DGX agent

arXiv:2604.11364v1 Announce Type: new Abstract: The two most influential cognitive architecture frameworks for AI agents, CoALA [21] and JEPA [12], both lack an explicit Knowledge layer with its own p

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

The Phase Is the Gradient: Equilibrium Propagation for Frequency Learning in Kuramoto Networks

DGX agent

arXiv:2604.10272v1 Announce Type: new Abstract: We prove that in a coupled Kuramoto oscillator network at stable equilibrium, the physical phase displacement under weak output nudging is the gradient

model-releasesarxiv-cs-lg
14 Apr 2026
Model Releases

The Rise and Fall of G in AGI

DGX agent

arXiv:2604.09911v1 Announce Type: cross Abstract: In the psychological literature the term `general intelligence' describes correlations between abilities and not simply the number of abilities. This

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

The Salami Slicing Threat: Exploiting Cumulative Risks in LLM Systems

DGX agent

arXiv:2604.11309v1 Announce Type: cross Abstract: Large Language Models (LLMs) face prominent security risks from jailbreaking, a practice that manipulates models to bypass built-in security constrain

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

THEIA: Learning Complete Kleene Three-Valued Logic in a Pure-Neural Modular Architecture

DGX agent

arXiv:2604.11284v1 Announce Type: cross Abstract: We present THEIA, a modular neural architecture that learns complete Kleene three-valued logic (K3) end-to-end without any external symbolic solver, a

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

There are multiple realities of AI right now. And what you have access to drastically changes your workflows, trust in AI, and ability to ad…

DGX agent

There are multiple realities of AI right now. And what you have access to drastically changes your workflows, trust in AI, and ability to adapt to the future. Here’s the briefest state of the AI world

model-releasesallie-k--miller--x
14 Apr 2026
Model Releases

Think in Sentences: Explicit Sentence Boundaries Enhance Language Model's Capabilities

DGX agent

arXiv:2604.10135v1 Announce Type: cross Abstract: Researchers have explored different ways to improve large language models (LLMs)' capabilities via dummy token insertion in contexts. However, existin

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Think Parallax: Solving Multi-Hop Problems via Multi-View Knowledge-Graph-Based Retrieval-Augmented Generation

DGX agent

arXiv:2510.15552v3 Announce Type: replace-cross Abstract: Large language models (LLMs) still struggle with multi-hop reasoning over knowledge-graphs (KGs), and we identify a previously overlooked stru

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Thinking Fast, Thinking Wrong: Intuitiveness Modulates LLM Counterfactual Reasoning in Policy Evaluation

DGX agent

arXiv:2604.10511v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used for causal and counterfactual reasoning, yet their reliability in real-world policy evaluation remain

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

This is why we released liteparse :) Free, open-source, designed for agents. Natively supports OCR / screenshotting for deeper visual unders…

DGX agent

This is why we released liteparse :) Free, open-source, designed for agents. Natively supports OCR / screenshotting for deeper visual understanding in a document when needed. @kepano I just tried it t

model-releasesjerry-liu--x
14 Apr 2026
Model Releases

Three Roles, One Model: Role Orchestration at Inference Time to Close the Performance Gap Between Small and Large Agents

DGX agent

arXiv:2604.11465v1 Announce Type: new Abstract: Large language model (LLM) agents show promise on realistic tool-use tasks, but deploying capable agents on modest hardware remains challenging. We stud

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Time is Not a Label: Continuous Phase Rotation for Temporal Knowledge Graphs and Agentic Memory

DGX agent

arXiv:2604.11544v1 Announce Type: cross Abstract: Structured memory representations such as knowledge graphs are central to autonomous agents and other long-lived systems. However, most existing appro

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

TimeSeriesExamAgent: Creating Time Series Reasoning Benchmarks at Scale

DGX agent

arXiv:2604.10291v1 Announce Type: new Abstract: Large Language Models (LLMs) have shown promising performance in time series modeling tasks, but do they truly understand time series data? While multip

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

TinyGaze: Lightweight Gaze-Gesture Recognition on Commodity Mobile Devices

DGX agent

arXiv:2604.09658v1 Announce Type: cross Abstract: Gaze gestures can provide hands free input on mobile devices, but practical use requires (i) gestures users can learn and recall and (ii) recognition

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

Today we're launching a rebuilt version of Claude Code on desktop. The app has been redesigned for the ground up to make it easier than ever…

DGX agent

Today we're launching a rebuilt version of Claude Code on desktop. The app has been redesigned for the ground up to make it easier than ever to parallelize work with Claude. I haven't opened an IDE or

model-releasesthariq--x
14 Apr 2026
Model Releases

Token-Budget-Aware Pool Routing for Cost-Efficient LLM Inference

DGX agent

arXiv:2604.09613v1 Announce Type: cross Abstract: Production vLLM fleets provision every instance for worst-case context length, wasting 4-8x concurrency on the 80-95% of requests that are short and s

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Too Nice to Tell the Truth: Quantifying Agreeableness-Driven Sycophancy in Role-Playing Language Models

DGX agent

arXiv:2604.10733v1 Announce Type: cross Abstract: Large language models increasingly serve as conversational agents that adopt personas and role-play characters at user request. This capability, while

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Topo-ADV: Generating Topology-Driven Imperceptible Adversarial Point Clouds

DGX agent

arXiv:2604.09879v1 Announce Type: new Abstract: Deep neural networks for 3D point cloud understanding have achieved remarkable success in object classification and recognition, yet recent work shows t

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

TorchUMM: A Unified Multimodal Model Codebase for Evaluation, Analysis, and Post-training

DGX agent

arXiv:2604.10784v1 Announce Type: new Abstract: Recent advances in unified multimodal models (UMMs) have led to a proliferation of architectures capable of understanding, generating, and editing acros

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Toward Generalized Cross-Lingual Hateful Language Detection with Web-Scale Data and Ensemble LLM Annotations

DGX agent

arXiv:2604.09625v1 Announce Type: new Abstract: We study whether large-scale unlabelled web data and LLM-based synthetic annotations can improve multilingual hate speech detection. Starting from texts

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

Towards Multi-Source Domain Generalization for Sleep Staging with Noisy Labels

DGX agent

arXiv:2604.10009v1 Announce Type: cross Abstract: Automatic sleep staging is a multimodal learning problem involving heterogeneous physiological signals such as EEG and EOG, which often suffer from do

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

Tracing the Roots: A Multi-Agent Framework for Uncovering Data Lineage in Post-Training LLMs

DGX agent

arXiv:2604.10480v1 Announce Type: new Abstract: Post-training data plays a pivotal role in shaping the capabilities of Large Language Models (LLMs), yet datasets are often treated as isolated artifact

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Tracking High-order Evolutions via Cascading Low-rank Fitting

DGX agent

arXiv:2604.10980v1 Announce Type: new Abstract: Diffusion models have become the de facto standard for modern visual generation, including well-established frameworks such as latent diffusion and flow

model-releasesarxiv-cs-lg
14 Apr 2026
Model Releases

Training-Free Object-Background Compositional T2I via Dynamic Spatial Guidance and Multi-Path Pruning

DGX agent

arXiv:2604.09850v1 Announce Type: new Abstract: Existing text-to-image diffusion models, while excelling at subject synthesis, exhibit a persistent foreground bias that treats the background as a pass

model-releasesarxiv-cs-cv
14 Apr 2026
← Previous
1…445446447448449…465
Next →