AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,630Total entries
1Added by human
84,629Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,288 results
Model Releases

ReSAE: Residualized Sparse Autoencoders for Multi-Layer Transformer Interventions

DGX agent

arXiv:2605.27819v1 Announce Type: cross Abstract: Sparse autoencoders are usually trained one layer at a time, even though transformer residual stream activations are strongly coupled across depth. Th

model-releasesarxiv-cs-ai
28 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Resolution-free neural surrogates for geometric parameterization and mapping with spatially varying fields

DGX agent

arXiv:2605.28551v1 Announce Type: new Abstract: Many imaging problems require computing spatial transformations induced by spatially varying intensity, feature, or density fields. Canonical examples i

model-releasesarxiv-cs-cv
28 May 2026
Model Releases

Resource-Constrained Affect Modelling via Variance Regularisation Pruning

DGX agent

arXiv:2605.27479v1 Announce Type: cross Abstract: Affective computing systems are increasingly embedded in pervasive and interactive environments, such as adaptive games, assistive technologies, and r

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Revisiting 2D Foundation Models for Scalable 3D Medical Image Classification

DGX agent

arXiv:2512.12887v3 Announce Type: replace Abstract: 3D medical image classification is essential for modern clinical workflows. Medical foundation models (FMs) have emerged as a promising approach for

model-releasesarxiv-cs-cv
28 May 2026
Model Releases

Revisiting Metafeatures to Explain Model Differences on Tabular Data

DGX agent

arXiv:2605.28418v1 Announce Type: new Abstract: With the rise of tabular foundation models alongside traditional models still performing well on many tasks, choosing the right model for a tabular data

model-releasesarxiv-cs-lg
28 May 2026
Model Releases

RGC: a radio AGN classifier based on deep learning. I. A semi-supervised multiclass model for VLA images

DGX agent

arXiv:2510.22190v2 Announce Type: replace-cross Abstract: Bent radio active galactic nuclei (RAGNs) -- wide-angle tails (WATs) and narrow-angle tails (NATs) -- trace dense environments in galaxy group

model-releasesarxiv-cs-lg
28 May 2026
Model Releases

RMPL: Relation-aware Multi-task Progressive Learning with Stage-wise Training for Multimedia Event Extraction

DGX agent

arXiv:2602.13748v2 Announce Type: replace Abstract: Multimedia Event Extraction (MEE) aims to identify events and their arguments from documents that contain both text and images. It requires groundin

model-releasesarxiv-cs-cl
28 May 2026
Model Releases

Robust Moment-Based Estimation via Spectral Gradient Reweighting

DGX agent

arXiv:2605.27718v1 Announce Type: cross Abstract: Moment-based estimation is a theoretically attractive approach to parametric inference, especially when likelihood-based estimation is unavailable, mi

model-releasesarxiv-cs-lg
28 May 2026
Model Releases

RW-TTT: Batched Serving for Request-Owned Test-Time Training State

DGX agent

arXiv:2605.28053v1 Announce Type: new Abstract: Test-time training (TTT) adapts an LLM during generation by reading and updating request-owned state, such as fast weights, low-rank deltas, or streamin

model-releasesarxiv-cs-lg
28 May 2026
Model Releases

Safe In-Context Reinforcement Learning

DGX agent

arXiv:2509.25582v3 Announce Type: replace Abstract: In-context reinforcement learning (ICRL) is an emerging RL paradigm where an agent, after pretraining, can adapt to out-of-distribution test tasks w

model-releasesarxiv-cs-lg
28 May 2026
Model Releases

SAM-Enhanced Segmentation on Road Datasets: Balancing Critical Classes in Autonomous Driving

DGX agent

arXiv:2605.28136v1 Announce Type: new Abstract: Dense semantic segmentation is essential for autonomous driving, yet many multi-modal datasets lack pixel-level annotations. The Zenseact Open Dataset (

model-releasesarxiv-cs-cv
28 May 2026
Model Releases

SAME: Stabilized Mixture-of-Experts for Multimodal Continual Instruction Tuning

DGX agent

arXiv:2602.01990v2 Announce Type: replace-cross Abstract: Multimodal Large Language Models (MLLMs) achieve strong performance through instruction tuning, but real-world deployment requires them to con

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

SeeGroup: Multi-Layer Depth Estimation of Transparent Surfaces via Self-Determined Grouping

DGX agent

arXiv:2605.28735v1 Announce Type: new Abstract: Transparent objects are common in daily life, and it is important to understand their multilayer depth, including the transparent surface and the object

model-releasesarxiv-cs-cv
28 May 2026
Model Releases

Self-Supervised Online Robot-Agnostic Traversability Estimation for Open-World Environments

DGX agent

arXiv:2605.28442v1 Announce Type: cross Abstract: Self-supervised online traversability estimation enables robots to continuously learn from unlabeled open-world experiences and adapt their navigation

model-releasesarxiv-cs-cv
28 May 2026
Model Releases

SIGMA: Bridging Structural and Distributional Gaps for Vision Foundation Model Adaptation

DGX agent

arXiv:2605.27893v1 Announce Type: new Abstract: Vision Foundation Models (VFMs) have demonstrated impressive representational capabilities. However, adapting them to downstream tasks via full fine-tun

model-releasesarxiv-cs-cv
28 May 2026
Model Releases

Sign-Aware Gated Sparse Autoencoders: Modeling Anticorrelated Features with Bi-Jump-ReLU Activations

DGX agent

arXiv:2605.28149v1 Announce Type: new Abstract: Sparse Autoencoders (SAEs) extract interpretable features from Large Language Models, but standard variants enforce non-negativity, forcing separate lat

model-releasesarxiv-cs-lg
28 May 2026
Model Releases

Simorgh at SemEval-2026 task 7: Region-Aware Hybrid Retrieval for Low-Resource Cultural Reasoning in Multilingual Question Answering

DGX agent

arXiv:2605.27636v1 Announce Type: new Abstract: Although Large Language Models (LLMs) demonstrate excellent capabilities and performance for general reasoning tasks within the general public domain, t

model-releasesarxiv-cs-cl
28 May 2026
Model Releases

SkillGrad: Optimizing Agent Skills Like Gradient Descent

DGX agent

arXiv:2605.27760v1 Announce Type: new Abstract: Agent skills provide a lightweight way to adapt LLM agents to specialized domains by storing reusable procedural knowledge in structured files. However,

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

SmartIterator: Visual Analytics Workflows for Supervising Unsupervised Data Grouping

DGX agent

arXiv:2605.28219v1 Announce Type: cross Abstract: Unsupervised learning methods -- topic modeling, partition-based and density-based clustering -- produce data groupings without human guidance, yet ch

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

SNARE: Adaptive Scenario Synthesis for Eliciting Overeager Behavior in Coding Agents

DGX agent

arXiv:2605.28122v1 Announce Type: cross Abstract: A coding agent executes a benign task as a sequence of shell, file, and network actions, any of which can quietly exceed the authorized scope while th

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Snippet-Driven Supply Chain Discovery with LLMs: Scaling Visibility in China

DGX agent

arXiv:2605.27845v1 Announce Type: cross Abstract: Financial and economic research often relies on structured supply-chain disclosures and commercial databases. In China, supplier--customer disclosure

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Snowveil: A Framework for Decentralised Preference Discovery

DGX agent

arXiv:2512.18444v2 Announce Type: replace-cross Abstract: Aggregating subjective preferences in social choice traditionally assumes a trusted central authority. In contrast, this paper formalises Dece

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

SONIC-O1: A Real-World Benchmark for Evaluating Multimodal Large Language Models on Audio-Video Understanding

DGX agent

arXiv:2601.21666v2 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) are a major focus of recent AI research. However, most prior work focuses on static image understanding, wh

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Soro: A Lightweight Foundation Model and Chatbot for Tajik

DGX agent

arXiv:2605.27379v1 Announce Type: new Abstract: We present Soro, a family of Tajik-specialized conversational large language models (LLMs) designed for real-world deployment under tight compute and co

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Sparse POD Mode Selection and Manifold Dimensionality Reduction with Neural Networks

DGX agent

arXiv:2605.27756v1 Announce Type: cross Abstract: High-performance computing enables simulation of high-dimensional physical systems, but downstream analyses such as inverse problems and control remai

model-releasesarxiv-cs-lg
28 May 2026
Model Releases

SSR3D-LLM: Structured Spatial Reasoning via Latent Steps for Fine-Grained Grounding in Unified 3D-LLMs

DGX agent

arXiv:2605.28490v1 Announce Type: cross Abstract: 3D object grounding localizes referred objects in a 3D scene from natural language. Unified instance-centric 3D-LLMs aim to solve grounding together w

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Stay Fair! Ensuring Group Fairness in Diffusion Models Across Guidance Scales

DGX agent

arXiv:2605.28036v1 Announce Type: new Abstract: Diffusion models steer conditional generation with a tunable guidance scale to trade off prompt alignment and diversity. However, existing debiasing tec

model-releasesarxiv-cs-cv
28 May 2026
Model Releases

Stochastic Gradient Descent with Momentum is Algorithmically Stable

DGX agent

arXiv:2605.28517v1 Announce Type: cross Abstract: Stochastic gradient descent with momentum (SGDM) is one of the most widely used optimization algorithms in machine learning. While optimization proper

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

StoryLens: Preference-Aligned Story Rewriting via Context-Aware Narrative Enrichment

DGX agent

arXiv:2605.28073v1 Announce Type: cross Abstract: Story rewriting aims to adapt existing narratives to diverse reader preferences while preserving plot consistency and narrative coherence. Unlike conv

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

StoryMI: Steerable Multi-Agent Therapeutic Dialogue Generation

DGX agent

arXiv:2605.27393v1 Announce Type: cross Abstract: Large language models (LLMs) can generate fluent dialogue, but prior works lack situational grounding, dynamic strategy control, and evaluation aligne

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

STR Robot: Design of an Autonomous Mobile Robot from Simulation to Reality

DGX agent

arXiv:2605.28110v1 Announce Type: new Abstract: With the rapid development of simulation tools, the development and validation of autonomous robotic systems have become more efficient before real-worl

model-releasesarxiv-cs-ro
28 May 2026
Model Releases

Structured Belief State and the First Precision-Aware Benchmark for LLM Memory Retrieval

DGX agent

arXiv:2605.11325v2 Announce Type: replace-cross Abstract: Every major benchmark for LLM memory systems, LoCoMo foremost, measures whether a model answered correctly, not whether the memory system retr

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

SuiChat-CN: Benchmarking Contextual Suicide Risk Assessment in Chinese Group Chats

DGX agent

arXiv:2605.27911v1 Announce Type: new Abstract: Suicide is a critical global public health challenge, causing approximately 720,000 deaths each year and calling for timely, effective prevention strate

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

SuperValid: Capability-Aligned OOD Validation for Generalizable Downstream Scaling

DGX agent

arXiv:2605.28179v1 Announce Type: new Abstract: Scaling laws guide large language model training by relating compute to cross-entropy loss, and recent work further extends them to predict downstream b

model-releasesarxiv-cs-cl
28 May 2026
Model Releases

Tabero: Learning Gentle Manipulation with Closed-Loop Force Feedback from Vision, Touch, and Language

DGX agent

arXiv:2605.27886v1 Announce Type: new Abstract: Tactile sensing is essential for robots to achieve human-like gentle manipulation. However, existing Vision-Language-Action (VLA) models struggle to exp

model-releasesarxiv-cs-ro
28 May 2026
Model Releases

Tackling Multimodal Learning Challenges with Mixture-of-Expert: A Survey

DGX agent

arXiv:2605.27431v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) presents a naturally compatible and scalable framework for multimodal learning, demonstrating strong adaptability across dive

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

TCP-MCP: Landscape-Guided Co-Evolution of Prompts and Communication Topologies for Multi-Agent Systems

DGX agent

arXiv:2605.27850v1 Announce Type: new Abstract: Effective multi-agent systems cannot be designed by selecting prompts or communication graphs in isolation. Agent behavior depends on the information an

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

The Abstraction Gap in Vision-Language Causal Reasoning

DGX agent

arXiv:2605.28779v1 Announce Type: new Abstract: Vision-language models (VLMs) generate fluent causal explanations, but current evaluations cannot distinguish linguistic plausibility from faithful caus

model-releasesarxiv-cs-cl
28 May 2026
Model Releases

The Alignment Floor: When Persona Customization Is Safe

DGX agent

arXiv:2605.27382v1 Announce Type: cross Abstract: A key promise of pluralistic AI is behavioral adaptation: persona prompts like 'be creative' or 'be thorough' let systems respect diverse user values

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

The Cases LJP Never Sees: Prosecution Decision Prediction for More Complete Criminal Liability Assessment

DGX agent

arXiv:2605.28464v1 Announce Type: cross Abstract: Legal Judgment Prediction (LJP) has become a core benchmark for evaluating AI in the criminal legal domain, but it only sees criminal cases that have

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

The Fragility of Chain-of-Thought Monitoring Across Typologically Diverse Languages

DGX agent

arXiv:2605.27901v1 Announce Type: cross Abstract: Chain-of-thought (CoT) monitoring has been proposed as a promising safety mechanism for detecting misaligned behavior in large language models. Howeve

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

The Harder Text Embedding Benchmark (HTEB): Beyond One-dimensional Static Robustness

DGX agent

arXiv:2605.28190v1 Announce Type: new Abstract: Embedding benchmarks like MTEB report a single score per model, implicitly treating robustness as a static, scalar property. We argue that embedding rob

model-releasesarxiv-cs-cl
28 May 2026
Model Releases

The Importance of Being Statistically Earnest: A Critical Re-evaluation of GSM-Symbolic

DGX agent

arXiv:2605.28700v1 Announce Type: new Abstract: The GSM-Symbolic benchmark (Mirzadeh et al., 2025) reported consistent performance drops across 25 Large Language Models (LLMs) when tested on template-

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

The Missing Piece in Pre-trained Model Evaluation: Reward-Guided Decoding Unlocks Task-Oriented Behavior Without Parameter Updates

DGX agent

arXiv:2605.28020v1 Announce Type: new Abstract: With the rapid progress of large language models (LLMs), reliably evaluating the capabilities of pre-trained LLMs has become increasingly important. The

model-releasesarxiv-cs-cl
28 May 2026
Model Releases

The Point, the Vision and the Text: Does Point Cloud Boost Spatial Reasoning of Large Language Models? A Bias-Controlled Study

DGX agent

arXiv:2504.04540v2 Announce Type: replace-cross Abstract: 3D Large Language Models (LLMs) leveraging spatial information in point clouds for 3D spatial reasoning attract great attention. Despite some

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

The Script is All You Need: An Agentic Framework for Long-Horizon Dialogue-to-Cinematic Video Generation

DGX agent

arXiv:2601.17737v3 Announce Type: replace-cross Abstract: Recent advances in video generation have produced models capable of synthesizing stunning visual content from simple text prompts. However, th

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Thermodynamic properties of chemically disordered compounds via AI-driven estimation of partition function with the PULSE method

DGX agent

arXiv:2605.28594v1 Announce Type: cross Abstract: In this article, we present an improved version of the PULSE method (Partition function Unsupervised Learning Sampling and Evaluation) for estimating

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Tool Forge: A Validation-Carrying Toolchain for Governed Agentic Execution

DGX agent

arXiv:2605.28000v1 Announce Type: cross Abstract: Large language model agents are increasingly expected to perform operational work: calling APIs, manipulating files, assembling workflows, and acting

model-releasesarxiv-cs-ai
28 May 2026
← Previous
1…190191192193194…361
Next →