AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
16,965 results
Model Releases

Shared Emotion Geometry Across Small Language Models: A Cross-Architecture Study of Representation, Behavior, and Methodological Confounds

DGX agent

arXiv:2604.11050v1 Announce Type: cross Abstract: We extract 21-emotion vector sets from twelve small language models (six architectures x base/instruct, 1B-8B parameters) under a unified comprehensio

model-releasesarxiv-cs-ai
14 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Sign Language Recognition in the Age of LLMs

DGX agent

arXiv:2604.11225v1 Announce Type: cross Abstract: Recent Vision Language Models (VLMs) have demonstrated strong performance across a wide range of multimodal reasoning tasks. This raises the question

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

SignReasoner: Compositional Reasoning for Complex Traffic Sign Understanding via Functional Structure Units

DGX agent

arXiv:2604.10436v1 Announce Type: new Abstract: Accurate semantic understanding of complex traffic signs-including those with intricate layouts, multi-lingual text, and composite symbols-is critical f

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

SimBench: Benchmarking the Ability of Large Language Models to Simulate Human Behaviors

DGX agent

arXiv:2510.17516v4 Announce Type: replace-cross Abstract: Large language model (LLM) simulations of human behavior have the potential to revolutionize the social and behavioral sciences, if and only i

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Simple but Stable, Fast and Safe: Achieve End-to-end Control by High-Fidelity Differentiable Simulation

DGX agent

arXiv:2604.10548v1 Announce Type: new Abstract: Obstacle avoidance is a fundamental vision-based task essential for enabling quadrotors to perform advanced applications. When planning the trajectory,

model-releasesarxiv-cs-ro
14 Apr 2026
Model Releases

Simulating Organized Group Behavior: New Framework, Benchmark, and Analysis

DGX agent

arXiv:2604.09874v1 Announce Type: new Abstract: Simulating how organized groups (e.g., corporations) make decisions (e.g., responding to a competitor's move) is essential for understanding real-world

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

Simulator Adaptation for Sim-to-Real Learning of Legged Locomotion via Proprioceptive Distribution Matching

DGX agent

arXiv:2604.11090v1 Announce Type: new Abstract: Simulation trained legged locomotion policies often exhibit performance loss on hardware due to dynamics discrepancies between the simulator and the rea

model-releasesarxiv-cs-ro
14 Apr 2026
Model Releases

Single-Agent LLMs Outperform Multi-Agent Systems on Multi-Hop Reasoning Under Equal Thinking Token Budgets

DGX agent

arXiv:2604.02460v2 Announce Type: replace Abstract: Recent work reports strong performance from multi-agent LLM systems (MAS), but these gains are often confounded by increased test-time computation.

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

SLM Finetuning for Natural Language to Domain Specific Code Generation in Production

DGX agent

arXiv:2604.09952v1 Announce Type: new Abstract: Many applications today use large language models for code generation; however, production systems have strict latency requirements that can be difficul

model-releasesarxiv-cs-lg
14 Apr 2026
Model Releases

SMART: When is it Actually Worth Expanding a Speculative Tree?

DGX agent

arXiv:2604.09731v1 Announce Type: cross Abstract: Tree-based speculative decoding accelerates autoregressive generation by verifying a branching tree of draft tokens in a single target-model forward p

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

SMFormer: Empowering Self-supervised Stereo Matching via Foundation Models and Data Augmentation

DGX agent

arXiv:2604.10218v1 Announce Type: new Abstract: Recent self-supervised stereo matching methods have made significant progress. They typically rely on the photometric consistency assumption, which pres

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

SmileyLlama: Modifying Large Language Models for Directed Chemical Space Exploration

DGX agent

arXiv:2409.02231v5 Announce Type: replace-cross Abstract: We show that large language model (LLMs) can be transformed via supervised fine-tuning (SFT) of engineered prompts into SmileyLlama for explor

model-releasesarxiv-cs-lg
14 Apr 2026
Model Releases

SODA: Semi On-Policy Black-Box Distillation for Large Language Models

DGX agent

arXiv:2604.03873v2 Announce Type: replace-cross Abstract: Black-box knowledge distillation for large language models presents a strict trade-off. Simple off-policy methods (e.g., sequence-level knowle

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

Solving Physics Olympiad via Reinforcement Learning on Physics Simulators

DGX agent

arXiv:2604.11805v1 Announce Type: cross Abstract: We have witnessed remarkable advances in LLM reasoning capabilities with the advent of DeepSeek-R1. However, much of this progress has been fueled by

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Spatial Competence Benchmark

DGX agent

arXiv:2604.09594v1 Announce Type: new Abstract: Spatial competence is the quality of maintaining a consistent internal representation of an environment and using it to infer discrete structure and pla

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

SpatialScore: Towards Comprehensive Evaluation for Spatial Intelligence

DGX agent

arXiv:2505.17012v3 Announce Type: replace-cross Abstract: Existing evaluations of multimodal large language models (MLLMs) on spatial intelligence are typically fragmented and limited in scope. In thi

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

SpecMoE: A Fast and Efficient Mixture-of-Experts Inference via Self-Assisted Speculative Decoding

DGX agent

arXiv:2604.10152v1 Announce Type: new Abstract: The Mixture-of-Experts (MoE) architecture has emerged as a promising approach to mitigate the rising computational costs of large language models (LLMs)

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

SPEED-Bench: A Unified and Diverse Benchmark for Speculative Decoding

DGX agent

arXiv:2604.09557v1 Announce Type: cross Abstract: Speculative Decoding (SD) has emerged as a critical technique for accelerating Large Language Model (LLM) inference. Unlike deterministic system optim

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Spoiler Alert: Narrative Forecasting as a Metric for Tension in LLM Storytelling

DGX agent

arXiv:2604.09854v1 Announce Type: new Abstract: LLMs have so far failed both to generate consistently compelling stories and to recognize this failure--on the leading creative-writing benchmark (EQ-Be

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

SRBench: A Comprehensive Benchmark for Sequential Recommendation with Large Language Models

DGX agent

arXiv:2604.09553v1 Announce Type: cross Abstract: LLM development has aroused great interest in Sequential Recommendation (SR) applications. However, comprehensive evaluation of SR models remains lack

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

StableTTA: Training-Free Test-Time Adaptation that Improves Model Accuracy on ImageNet1K to 96%

DGX agent

arXiv:2604.04552v2 Announce Type: replace-cross Abstract: Ensemble methods are widely used to improve predictive performance, but their effectiveness often comes at the cost of increased memory usage

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

STaR-DRO: Stateful Tsallis Reweighting for Group-Robust Structured Prediction

DGX agent

arXiv:2604.09737v1 Announce Type: cross Abstract: Structured prediction requires models to generate ontology-constrained labels, grounded evidence, and valid structure under ambiguity, label skew, and

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

STARS: Skill-Triggered Audit for Request-Conditioned Invocation Safety in Agent Systems

DGX agent

arXiv:2604.10286v1 Announce Type: new Abstract: Autonomous language-model agents increasingly rely on installable skills and tools to complete user tasks. Static skill auditing can expose capability s

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

StarVLA-alpha: Reducing Complexity in Vision-Language-Action Systems

DGX agent

arXiv:2604.11757v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have recently emerged as a promising paradigm for building general-purpose robotic agents. However, the VLA landsc

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

STORM: End-to-End Referring Multi-Object Tracking in Videos

DGX agent

arXiv:2604.10527v1 Announce Type: cross Abstract: Referring multi-object tracking (RMOT) is a task of associating all the objects in a video that semantically match with given textual queries or refer

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

StyleBench: Evaluating thinking styles in Large Language Models

DGX agent

arXiv:2509.20868v2 Announce Type: replace-cross Abstract: Structured reasoning can improve the inference performance of large language models (LLMs), but it also introduces computational cost and cont

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Switch-JustDance: Benchmarking Whole Body Motion Tracking Controllers Using a Commercial Console Game

DGX agent

arXiv:2511.17925v3 Announce Type: replace-cross Abstract: Recent advances in whole-body robot control have enabled humanoid and legged robots to perform increasingly agile and coordinated motions. How

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

SymphoMotion: Joint Control of Camera Motion and Object Dynamics for Coherent Video Generation

DGX agent

arXiv:2604.03723v2 Announce Type: replace Abstract: Controlling both camera motion and object dynamics is essential for coherent and expressive video generation, yet current methods typically handle o

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

Synthius-Mem: Brain-Inspired Hallucination-Resistant Persona Memory Achieving 94.4% Memory Accuracy and 99.6% Adversarial Robustness on LoCoMo

DGX agent

arXiv:2604.11563v1 Announce Type: cross Abstract: Providing AI agents with reliable long-term memory that does not hallucinate remains an open problem. Current approaches to memory for LLM agents -- s

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

TAG-Head: Time-Aligned Graph Head for Plug-and-Play Fine-grained Action Recognition

DGX agent

arXiv:2604.11498v1 Announce Type: new Abstract: Fine-grained human action recognition (FHAR) is challenging because visually similar actions differ by subtle spatio-temporal cues. Many recent systems

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

Tail-Aware Information-Theoretic Generalization for RLHF and SGLD

DGX agent

arXiv:2604.10727v1 Announce Type: cross Abstract: Classical information-theoretic generalization bounds typically control the generalization gap through KL-based mutual information and therefore rely

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Teaching Language Models How to Code Like Learners: Conversational Serialization for Student Simulation

DGX agent

arXiv:2604.10720v1 Announce Type: new Abstract: Artificial models that simulate how learners act and respond within educational systems are a promising tool for evaluating tutoring strategies and feed

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Template-assisted Contrastive Learning of Task-oriented Dialogue Sentence Embeddings

DGX agent

arXiv:2305.14299v3 Announce Type: replace-cross Abstract: Learning high quality sentence embeddings from dialogues has drawn increasing attentions as it is essential to solve a variety of dialogue-ori

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

TempusBench: An Evaluation Framework for Time-Series Forecasting

DGX agent

arXiv:2604.11529v1 Announce Type: new Abstract: Foundation models have transformed natural language processing and computer vision, and a rapidly growing literature on time-series foundation models (T

model-releasesarxiv-cs-lg
14 Apr 2026
Model Releases

Text-to-Image Models and Their Representation of People from Different Nationalities Engaging in Activities

DGX agent

arXiv:2504.06313v5 Announce Type: replace Abstract: This paper investigates how popular text-to-image (T2I) models, DALL-E 3 and Gemini 3 Pro Preview, depict people from 206 nationalities when prompte

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

The Amazing Agent Race: Strong Tool Users, Weak Navigators

DGX agent

arXiv:2604.10261v1 Announce Type: new Abstract: Existing tool-use benchmarks for LLM agents are overwhelmingly linear: our analysis of six benchmarks shows 55 to 100% of instances are simple chains of

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

The Blind Spot of Agent Safety: How Benign User Instructions Expose Critical Vulnerabilities in Computer-Use Agents

DGX agent

arXiv:2604.10577v1 Announce Type: cross Abstract: Computer-use agents (CUAs) can now autonomously complete complex tasks in real digital environments, but when misled, they can also be used to automat

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

The Missing Knowledge Layer in Cognitive Architectures for AI Agents

DGX agent

arXiv:2604.11364v1 Announce Type: new Abstract: The two most influential cognitive architecture frameworks for AI agents, CoALA [21] and JEPA [12], both lack an explicit Knowledge layer with its own p

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

The Phase Is the Gradient: Equilibrium Propagation for Frequency Learning in Kuramoto Networks

DGX agent

arXiv:2604.10272v1 Announce Type: new Abstract: We prove that in a coupled Kuramoto oscillator network at stable equilibrium, the physical phase displacement under weak output nudging is the gradient

model-releasesarxiv-cs-lg
14 Apr 2026
Model Releases

The Rise and Fall of G in AGI

DGX agent

arXiv:2604.09911v1 Announce Type: cross Abstract: In the psychological literature the term `general intelligence' describes correlations between abilities and not simply the number of abilities. This

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

The Salami Slicing Threat: Exploiting Cumulative Risks in LLM Systems

DGX agent

arXiv:2604.11309v1 Announce Type: cross Abstract: Large Language Models (LLMs) face prominent security risks from jailbreaking, a practice that manipulates models to bypass built-in security constrain

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

THEIA: Learning Complete Kleene Three-Valued Logic in a Pure-Neural Modular Architecture

DGX agent

arXiv:2604.11284v1 Announce Type: cross Abstract: We present THEIA, a modular neural architecture that learns complete Kleene three-valued logic (K3) end-to-end without any external symbolic solver, a

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Think in Sentences: Explicit Sentence Boundaries Enhance Language Model's Capabilities

DGX agent

arXiv:2604.10135v1 Announce Type: cross Abstract: Researchers have explored different ways to improve large language models (LLMs)' capabilities via dummy token insertion in contexts. However, existin

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Think Parallax: Solving Multi-Hop Problems via Multi-View Knowledge-Graph-Based Retrieval-Augmented Generation

DGX agent

arXiv:2510.15552v3 Announce Type: replace-cross Abstract: Large language models (LLMs) still struggle with multi-hop reasoning over knowledge-graphs (KGs), and we identify a previously overlooked stru

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Thinking Fast, Thinking Wrong: Intuitiveness Modulates LLM Counterfactual Reasoning in Policy Evaluation

DGX agent

arXiv:2604.10511v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used for causal and counterfactual reasoning, yet their reliability in real-world policy evaluation remain

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Three Roles, One Model: Role Orchestration at Inference Time to Close the Performance Gap Between Small and Large Agents

DGX agent

arXiv:2604.11465v1 Announce Type: new Abstract: Large language model (LLM) agents show promise on realistic tool-use tasks, but deploying capable agents on modest hardware remains challenging. We stud

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Time is Not a Label: Continuous Phase Rotation for Temporal Knowledge Graphs and Agentic Memory

DGX agent

arXiv:2604.11544v1 Announce Type: cross Abstract: Structured memory representations such as knowledge graphs are central to autonomous agents and other long-lived systems. However, most existing appro

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

TimeSeriesExamAgent: Creating Time Series Reasoning Benchmarks at Scale

DGX agent

arXiv:2604.10291v1 Announce Type: new Abstract: Large Language Models (LLMs) have shown promising performance in time series modeling tasks, but do they truly understand time series data? While multip

model-releasesarxiv-cs-ai
14 Apr 2026
← Previous
1…340341342343344…354
Next →