AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,553 results
5 May 2026

Arithmetic in the Wild: Llama uses Base-10 Addition to Reason About Cyclic Concepts

Model ReleasesDGX agent

arXiv:2605.01148v1 Announce Type: cross Abstract: Does structure in representations imply structure in computation? We study how Llama-3.1-8B reasons over cyclic concepts (e.g., 'what month is six mon

ARMOR 2025: A Military-Aligned Benchmark for Evaluating Large Language Model Safety Beyond Civilian Contexts

Model ReleasesDGX agent

arXiv:2605.00245v1 Announce Type: new Abstract: Large language models (LLMs) are now being explored for defense applications that require reliable and legally compliant decision support. They also hol

Assistance Without Interruption: A Benchmark and LLM-based Framework for Non-Intrusive Human-Robot Assistance


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases
DGX agent

arXiv:2605.01368v1 Announce Type: new Abstract: Human-robot interaction (HRI) has long studied how agents and people coordinate to achieve shared goals. In this work, we formalize and benchmark the no

ATR-Bench: A Federated Learning Benchmark for Adaptation, Trust, and Reasoning

Model ReleasesDGX agent

arXiv:2505.16850v2 Announce Type: replace-cross Abstract: Federated Learning (FL) has emerged as a promising paradigm for collaborative model training while preserving data privacy across decentralize

Attention Is Where You Attack

Model ReleasesDGX agent

arXiv:2605.00236v1 Announce Type: cross Abstract: Safety-aligned large language models rely on RLHF and instruction tuning to refuse harmful requests, yet the internal mechanisms implementing safety b

AttnRouter: Per-Category Attention Routing for Training-Free Image Editing on MMDiT

Model ReleasesDGX agent

arXiv:2605.01480v1 Announce Type: new Abstract: We study training-free image editing on Qwen-Image-Edit-2511, a 60-block multi-modal diffusion transformer (MMDiT) that concatenates noise and source-im

Automated Interpretability and Feature Discovery in Language Models with Agents

Model ReleasesDGX agent

arXiv:2605.01555v1 Announce Type: new Abstract: We introduce an autonomous multiagent framework for mechanistic interpretability that automates both explaining and finding internal features in large l

AutoSpatial: Visual-Language Reasoning for Social Robot Navigation through Efficient Spatial Reasoning Learning

Model ReleasesDGX agent

arXiv:2503.07557v2 Announce Type: replace Abstract: We present a novel method, AutoSpatial, an efficient approach with structured spatial grounding to enhance VLMs' spatial reasoning. By combining min

AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models

Model ReleasesDGX agent

arXiv:2506.09082v5 Announce Type: replace Abstract: The rise of vision foundation models (VFMs) calls for systematic evaluation. A common approach pairs VFMs with large language models (LLMs) as gener

BadmintonGRF: A Multimodal Dataset and Benchmark for Markerless Ground Reaction Force Estimation in Badminton

Model ReleasesDGX agent

arXiv:2605.01876v1 Announce Type: new Abstract: Multimodal resources for non-periodic court sports with laboratory-grade sensing remain scarce: few publicly pair instrumented ground reaction force (GR

Beating the Style Detector: Three Hours of Agentic Research on the AI-Text Arms Race

Model ReleasesDGX agent

arXiv:2605.02620v1 Announce Type: new Abstract: Reproducing an empirical NLP study used to take weeks. Given the released data and a modern agentic-research harness, we redo every experiment of a rece

Benchmarking local Hebbian learning rules for memory storage and prototype extraction

Model ReleasesDGX agent

arXiv:2605.01074v1 Announce Type: cross Abstract: Associative memory or content-addressable memory is an important component function in computer science and information processing, and at the same ti

Benchmarking Retrieval Strategies for Biomedical Retrieval-Augmented Generation: A Controlled Empirical Study

Model ReleasesDGX agent

arXiv:2605.02520v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) offers a well-established path to grounding large language model (LLM) outputs in external knowledge, yet the quest

Benchmarking Single-Pose Docking, Consensus Rescoring, and Supervised ML on the LIT-PCBA Library: A Critical Evaluation of DiffDock, AutoDock-GPU, GNINA, and DiffDock-NMDN

Model ReleasesDGX agent

arXiv:2605.01681v1 Announce Type: new Abstract: Virtual screening performance depends heavily on the chosen docking and scoring methods. Recent AI-based tools such as DiffDock and NMDN have reported s

Benchmarking Wireless Representations: High-Dimensional vs. Compressed Embeddings for Efficiency and Robustness

Model ReleasesDGX agent

arXiv:2605.02009v1 Announce Type: cross Abstract: Building on recent advances in representation learning for wireless channels, this work investigates the cost-benefit trade-offs of high-dimensional c

Beyond Perplexity: Character Distribution Signatures and the MDTA Benchmark for AI Text Detection

Model ReleasesDGX agent

arXiv:2605.01647v1 Announce Type: new Abstract: Training-free AI text detection methods primarily rely on model log-probabilities, achieving strong performance through approaches like Binoculars and D

BIM Information Extraction Through LLM-based Adaptive Exploration

Model ReleasesDGX agent

arXiv:2605.01698v1 Announce Type: new Abstract: BIM models provide structured representations of building geometry, semantics, and topology, yet extracting specific information from them remains remar

Book publishers sue Meta over AI’s ‘word-for-word’ copying

Model ReleasesDGX agent

Meta is facing a class action lawsuit filed by five major book publishers and one author over claims the company 'engaged in one of the most massive infringements of copyrighted materials in history'

Boosting Multimodal Remote Sensing Image Classification with Transformer-based Heterogeneously Salient Graph Representation

Model ReleasesDGX agent

arXiv:2311.10320v3 Announce Type: replace Abstract: Data collected by different modalities can provide a wealth of complementary information, such as hyperspectral image (HSI) to offer rich spectral-s

Break the Block: Dynamic-size Reasoning Blocks for Diffusion Large Language Models via Monotonic Entropy Descent with Reinforcement Learning

Model ReleasesDGX agent

arXiv:2605.02263v1 Announce Type: new Abstract: Recent diffusion large language models (dLLMs) have demonstrated both effectiveness and efficiency in reasoning via a block-based semi-autoregressive ge

Breaking the Silence: A Dataset and Benchmark for Bangla Text-to-Gloss Translation

Model ReleasesDGX agent

arXiv:2504.02293v3 Announce Type: replace Abstract: Gloss is a written approximation that bridges Sign Language (SL) and its corresponding spoken language. Despite a deaf and hard-of-hearing populatio

BRITE: A Benchmark for Reliable and Interpretable T2V Evaluation on Implausible Scenarios

Model ReleasesDGX agent

arXiv:2605.00873v1 Announce Type: cross Abstract: The rapid advancement of photorealistic Text-to-Video (T2V) generation brings in an urgent need for up-to-date evaluation methods. Existing benchmarks

Can LLMs Compress (and Decompress)? Evaluating Code Understanding and Execution via Invertibility

Model ReleasesDGX agent

arXiv:2601.13398v2 Announce Type: replace Abstract: LLMs demonstrate strong performance on code benchmarks, yet consistent reasoning across forward and backward execution remains elusive. We present R

CASE: An Agentic AI Framework for Enhancing Scam Intelligence in Digital Payments

Model ReleasesDGX agent

arXiv:2508.19932v2 Announce Type: replace Abstract: The proliferation of digital payment platforms has transformed commerce, offering unmatched convenience and accessibility globally. However, this gr

Causal2Vec: Improving Decoder-only LLMs as Embedding Models through a Contextual Token

Model ReleasesDGX agent

arXiv:2507.23386v3 Announce Type: replace Abstract: Decoder-only large language models (LLMs) have been increasingly adopted to build embedding models for diverse tasks. To overcome the inherent limit

CEZSAR: A Contrastive Embedding Method for Zero-Shot Action Recognition

Model ReleasesDGX agent

arXiv:2605.01165v1 Announce Type: new Abstract: This paper proposes a novel Zero-Shot Action Recognition~(ZSAR) method based on contrastive learning. In ZSAR, we aim to classify examples from classes

CGFformer: Cluster-Guidance Frequency Transformer for Pansharpening

Model ReleasesDGX agent

arXiv:2605.01490v1 Announce Type: new Abstract: Pansharpening aims to generate high-resolution multispectral (HRMS) images by fusing low-resolution multispectral (LRMS) images with high-resolution pan

Channel-Level Relation to Attentive Aggregation with Neighborhood-Homogeneity Constraint for Point Cloud Analysis

Model ReleasesDGX agent

arXiv:2605.02357v1 Announce Type: new Abstract: In 3D point cloud understanding, the core challenge lies in accurately capturing discriminative features within complex neighborhoods, which directly af

Chart-FR1: Visual Focus-Driven Fine-Grained Reasoning on Dense Charts

Model ReleasesDGX agent

arXiv:2605.01882v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have shown considerable potential in chart understanding and reasoning tasks. However, they still struggle with

Chebyshev-Augmented One-Shot Transfer Learning for PINNs on Nonlinear Differential Equations

Model ReleasesDGX agent

arXiv:2605.01634v1 Announce Type: new Abstract: Physics-Informed Neural Networks (PINNs) offer a flexible paradigm for solving differential equations by embedding governing laws into the training obje

Checkerboard: A Simple, Effective, Efficient and Learning-free Clean Label Backdoor Attack with Low Poisoning Budget

Model ReleasesDGX agent

arXiv:2605.01298v1 Announce Type: cross Abstract: Backdoor attacks threaten the deep learning supply chain by poisoning a small fraction of the training data so that a model behaves normally on clean

CLaC at SemEval-2026 Task 6: Response Clarity Detection in Political Discourse

Model ReleasesDGX agent

arXiv:2605.02170v1 Announce Type: new Abstract: In this paper, we present our system for SemEval-2026 Task 6 (CLARITY) on response clarity and evasion detection in question-answer pairs from U.S. pres

Class-Aware Adaptive Differential Privacy in Deep Learning for Sensor-Based Fall Detection

Model ReleasesDGX agent

arXiv:2605.01679v1 Announce Type: cross Abstract: Fall detection is a critical task in healthcare, particularly for elderly people. Timely fall detection and treatment can prevent severe injuries. Sen

Cloud Engineer’s AI Toolkit: Sign up Now for a Developer Workshop Near You!

Model ReleasesDGX agent

The world of AI is rapidly shifting from experimental Large Language Models to an era of Agentic AI. In the agentic era, autonomous software agents act on behalf of employees and consumers—driving a f

CNN-based Multi-In-Multi-Out Model for Efficient Spatiotemporal Prediction

Model ReleasesDGX agent

arXiv:2605.01277v1 Announce Type: new Abstract: Recently, Convolutional Neural Network (CNN) or Transformer architecture based models have been proposed to overcome the limitations of Recurrent Neural

CoAction: Cross-task Correlation-aware Pareto Set Learning

Model ReleasesDGX agent

arXiv:2605.01712v1 Announce Type: new Abstract: Pareto set learning (PSL) is an emerging paradigm in multi-objective optimization that trains neural networks to map preference vectors to Pareto optima

COCORELI: Enforcing Execution Preconditions for Reliable Collaborative Instruction Following

Model ReleasesDGX agent

arXiv:2509.04470v2 Announce Type: replace Abstract: Autonomous agents executing human instructions must operate reliably even when instructions are incomplete. While recent approaches improve detectio

CombinationTS: A Modular Framework for Understanding Time-Series Forecasting Models

Model ReleasesDGX agent

arXiv:2605.01231v1 Announce Type: new Abstract: Recent progress in time-series forecasting has led to rapidly increasing architectural complexity, yet many reported State-of-the-Art gains are statisti

Component-Aware Self-Speculative Decoding in Hybrid Language Models

Model ReleasesDGX agent

arXiv:2605.01106v1 Announce Type: new Abstract: Speculative decoding accelerates autoregressive inference by drafting candidate tokens with a fast model and verifying them in parallel with the target.

Compute Optimal Tokenization

Model ReleasesDGX agent

arXiv:2605.01188v1 Announce Type: new Abstract: Scaling laws enable the optimal selection of data amount and language model size, yet the impact of the data unit, the token, on this relationship remai

Concepts Whisper While Syntax Shouts: Spectral Anti-Concentration and the Dual Geometry of Transformer Representations

Model ReleasesDGX agent

arXiv:2605.01609v1 Announce Type: new Abstract: We test whether the causal inner product of itet{park2024linear} -- defined by the unembedding covariance Sigma -- enables cross-lingual concept transpo

Congestion-Aware Dynamic Axonal Delay for Spiking Neural Networks

Model ReleasesDGX agent

arXiv:2605.01291v1 Announce Type: new Abstract: Spiking Neural Networks (SNNs) are widely regarded as an energy-efficient paradigm for modeling and processing temporal and event-driven information. In

Constructing Interpretable Features from Compositional Neuron Groups

Model ReleasesDGX agent

arXiv:2506.10920v2 Announce Type: replace Abstract: A central goal for mechanistic interpretability has been to identify the right units of analysis in large language models (LLMs) that causally expla

ContextualJailbreak: Evolutionary Red-Teaming via Simulated Conversational Priming

Model ReleasesDGX agent

arXiv:2605.02647v1 Announce Type: new Abstract: Large language models (LLMs) remain vulnerable to jailbreak attacks that bypass safety alignment and elicit harmful responses. A growing body of work sh

Control Reinforcement Learning: Interpretable Token-Level Steering of LLMs via Sparse Autoencoder Features

Model ReleasesDGX agent

arXiv:2602.10437v3 Announce Type: replace-cross Abstract: Sparse autoencoders (SAEs) decompose language model activations into interpretable features, but existing methods reveal only which features a

Controllable Logical Hypothesis Generation for Abductive Reasoning in Knowledge Graphs

Model ReleasesDGX agent

arXiv:2505.20948v3 Announce Type: replace Abstract: Abductive reasoning in knowledge graphs aims to generate plausible logical hypotheses from observed entities, with broad applications in areas such

Coopetition-Gym v1: A Formally Grounded Platform for Mixed-Motive Multi-Agent Reinforcement Learning under Strategic Coopetition

Model ReleasesDGX agent

arXiv:2605.02063v1 Announce Type: cross Abstract: We present Coopetition-Gym v1, a benchmark platform for mixed-motive multi-agent reinforcement learning under strategic coopetition. The platform comp

CorrSteer: Generation-Time LLM Steering via Correlated Sparse Autoencoder Features

Model ReleasesDGX agent

arXiv:2508.12535v3 Announce Type: replace Abstract: Sparse Autoencoders (SAEs) can extract interpretable features from large language models (LLMs) without supervision. However, their effectiveness in

CoSpaDi: Compressing LLMs via Calibration-Guided Sparse Dictionary Learning

Model ReleasesDGX agent

arXiv:2509.22075v5 Announce Type: replace Abstract: Post-training compression of large language models (LLMs) often relies on low-rank weight approximations that represent each column of the weight ma

CP-SynC: Multi-Agent Zero-Shot Constraint Modeling in MiniZinc with Synthesized Checkers

Model ReleasesDGX agent

arXiv:2605.01675v1 Announce Type: cross Abstract: Constraint Programming (CP) is a powerful paradigm for solving combinatorial problems, yet translating natural language problem descriptions into exec

Creating and Evaluating Figurative Language Dataset for Sindhi

Model ReleasesDGX agent

arXiv:2605.01323v1 Announce Type: new Abstract: In this article, we introduce SiNFluD, a novel benchmark dataset for Sindhi figurative language classification. We first collect raw text from various b

Cross-Language Learning within Arabic Script for Low-Resource HTR

Model ReleasesDGX agent

arXiv:2605.02089v1 Announce Type: new Abstract: Handwritten Text Recognition (HTR) under limited labeled data remains a challenging problem, particularly for Arabic-script languages. Although modern s

CyclicJudge: Mitigating Judge Bias Efficiently in LLM-based Evaluation

Model ReleasesDGX agent

arXiv:2603.01865v3 Announce Type: replace Abstract: LLM-as-judge evaluation has become standard practice for open-ended model assessment; however, judges exhibit systematic biases that cannot be avera

datasette-referrer-policy 0.1

Model ReleasesDGX agent

Release: datasette-referrer-policy 0.1 The OpenStreetMap tiles on the Datasette global-power-plants demo weren't displaying correctly. This turned out to be caused by two bugs. The first is that the C

🚀 Day-0 MTP support for Gemma4 now available at vLLM with ready-to-use docker image! ⚡️Enjoy up to 3x faster decoding performance to superc…

Model ReleasesDGX agent

🚀 Day-0 MTP support for Gemma4 now available at vLLM with ready-to-use docker image! ⚡️Enjoy up to 3x faster decoding performance to supercharge your development with zero quality degradation! Check o

DBLP: Phase-Aware Bounded-Loss Transport for Burst-Resilient Distributed ML Training

Model ReleasesDGX agent

arXiv:2605.01989v1 Announce Type: new Abstract: Distributed machine learning (ML) training has become a necessity with the prevalence of billion to trillion-parameter-scale models. While prior work ha

Decoding-Time Debiasing via Process Reward Models: From Controlled Fill-in to Open-Ended Generation

Model ReleasesDGX agent

arXiv:2605.02348v1 Announce Type: new Abstract: Large language models pick up social biases from the data they are trained on and carry those biases into downstream applications, often reinforcing ste

Decompose and Recompose: Reasoning New Skills from Existing Abilities for Cross-Task Robotic Manipulation

Model ReleasesDGX agent

arXiv:2605.01448v1 Announce Type: cross Abstract: Cross-task generalization is a core challenge in open-world robotic manipulation, and the key lies in extracting transferable manipulation knowledge f

Deep neural networks with Fisher vector encoding for medical image classification

Model ReleasesDGX agent

arXiv:2605.01667v1 Announce Type: new Abstract: Orderless encoding methods have shown to improve Convolutional Neural Networks (CNNs) for image classification in the context of limited availability of

Deep Time Series Models: A Comprehensive Survey and Benchmark

Model ReleasesDGX agent

arXiv:2407.13278v3 Announce Type: replace Abstract: Time series, characterized by a sequence of data points organized in a discrete-time order, are ubiquitous in real-world scenarios. Unlike other dat

← Previous
1…285286287288289…376
Next →