AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
All
84,661Total entries
1Added by human
84,660Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,620 results
Model Releases

Partition of Unity Neural Networks for Interpretable Classification with Explicit Class Regions

DGX agent

arXiv:2602.00511v2 Announce Type: replace Abstract: Despite their empirical success, neural network classifiers remain difficult to interpret. In softmax-based models, class regions are defined implic

model-releasesarxiv-cs-lg
26 May 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Partner-Aware Hierarchical Skill Discovery for Robust Human-AI Collaboration

DGX agent

arXiv:2605.24352v1 Announce Type: new Abstract: Multi-agent collaboration, especially in human-AI teaming, requires agents that can adapt to novel partners with diverse and dynamic behaviors. Conventi

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

PDEInvBench: A Comprehensive Dataset and Design Space Exploration of Neural Networks for PDE Inverse Problems

DGX agent

arXiv:2605.25353v1 Announce Type: new Abstract: Inverse problems in partial differential equations (PDEs) involve estimating the physical parameters of a system from observed spatiotemporal solution f

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

PEDESTRIANQA: A Benchmark for Vision-Language Models on Pedestrian Intention and Trajectory Prediction

DGX agent

arXiv:2605.24562v1 Announce Type: cross Abstract: Pedestrian intention and trajectory prediction are critical for the safe deployment of autonomous driving systems, directly influencing navigation dec

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

PennySynth: RAG-Driven Data Synthesis for Automated Quantum Code Generation

DGX agent

arXiv:2605.25572v1 Announce Type: cross Abstract: The growing complexity of quantum programming frameworks has exposed a critical limitation in existing large language model (LLM)-based code assistant

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Personalize-then-Store: Benchmarking and Learning Personalized Memory for Long-horizon Agents

DGX agent

arXiv:2605.25535v1 Announce Type: new Abstract: Existing large language model (LLM) based memory systems apply universal, static policies that overlook a fundamental reality: the contexts that are wor

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Personalized Federated Learning by Energy-Efficient UAV Communications

DGX agent

arXiv:2605.25212v1 Announce Type: new Abstract: Federated learning (FL) is an effective paradigm for enhancing the learning capability of edge devices while preserving data privacy. In geographically

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

Physen-Noise2Noise: Physics-Guided Self-Supervised Defocus Deblurring with Bias Correction under Low-Light Conditions

DGX agent

arXiv:2605.24590v1 Announce Type: cross Abstract: Low-light, long-exposure defocus deblurring remains a challenging problem due to the simultaneous presence of severe blur and complex biased noise. Ex

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

PiXTime: A Model for Federated Time Series Forecasting with Heterogeneous Data across Nodes

DGX agent

arXiv:2601.05613v2 Announce Type: replace-cross Abstract: While collaborative forecasting on distributed time series is highly desirable, directly pooling localized datasets is often impractical due t

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Polymorphism Is Rotation: Operational Mechanistic Interpretability from a Two-Layer Transformer to Pythia-70m

DGX agent

arXiv:2605.24577v1 Announce Type: cross Abstract: Independently trained transformers compute the same function in residual-stream bases that differ by a uniform random rotation on SO(d_{model}). We ca

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

PolySAE: Modeling Feature Interactions in Sparse Autoencoders via Polynomial Decoding

DGX agent

arXiv:2602.01322v2 Announce Type: replace-cross Abstract: Sparse autoencoders (SAEs) interpret neural network representations by decomposing activations into sparse combinations of dictionary atoms. H

model-releasesarxiv-cs-cl
26 May 2026
Model Releases

Position: AI for Science Should Treat Measurement-to-Dataset Pipelines as Inference Components

DGX agent

arXiv:2605.24558v1 Announce Type: new Abstract: AI for Science (AI4Science) workflows often treat the released dataset as a fixed interface to the underlying system. However, in domains relying on ind

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

Pragmatic Reasoning improves LLM Code Generation

DGX agent

arXiv:2502.15835v5 Announce Type: replace-cross Abstract: Pragmatic reasoning helps interlocutors infer intended meaning from ambiguous or underspecified messages by considering shared context and cou

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Quantifying the Impact of Translation Errors on Multilingual LLM Evaluation

DGX agent

arXiv:2605.24904v1 Announce Type: new Abstract: Machine-translated benchmarks are widely used to assess the multilingual capabilities of large language models (LLMs), yet translation errors in these b

model-releasesarxiv-cs-cl
26 May 2026
Model Releases

Quaternion Self-Attention with Shared Scores

DGX agent

arXiv:2605.24920v1 Announce Type: cross Abstract: Quaternion neural networks are parameter-efficient and model multidimensional dependencies by representing four related features as a single entity. H

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Querying structural and functional niches on spatial transcriptomics data

DGX agent

arXiv:2410.10652v4 Announce Type: replace-cross Abstract: Cells in multicellular organisms coordinate to form structural and functional niches. With spatial transcriptomics (ST) enabling gene expressi

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

QUEST: Training Frontier Deep Research Agents with Fully Synthetic Tasks

DGX agent

arXiv:2605.24218v1 Announce Type: new Abstract: Deep research agents extend the role of search engines from retrieving keyword-matched pages to synthesizing knowledge, fundamentally changing how human

model-releasesarxiv-cs-cl
26 May 2026
Model Releases

QUIET: A Multi-Blank Cascaded Story Cloze Benchmark for LLM Creative Generation Capability

DGX agent

arXiv:2605.25955v1 Announce Type: cross Abstract: Large language models (LLMs) face a dual challenge in creative capability evaluation: existing benchmarks (e.g., Story Cloze Test, HellaSwag) measure

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Qwen 3.7 Max is now supported in Hermes Agent

DGX agent

Nous Research has added support for Qwen 3.7 Max, a large language model, within their Hermes Agent framework. This integration enables users to leverage Qwen 3.7 Max's capabilities when building or d

model-releasesnous-research--x
26 May 2026
Model Releases

Qwen3.7 Max now available in Go - text only - 1M context - smartest model in the Qwen family to date

DGX agent

Alibaba's Qwen team has released Qwen3.7 Max, their most advanced model to date, now available for use with Go programming language support. The model features a 1 million token context window and is

model-releasesqwen--x
26 May 2026
Model Releases

ran my first benchmark this weekend (longmemeval) mostly to test activegraph, learned a lot! - this is a stepping stone to show the event ba…

DGX agent

ran my first benchmark this weekend (longmemeval) mostly to test activegraph, learned a lot! - this is a stepping stone to show the event based agent system works. the AI convinced me not to start wit

model-releasesyohei-nakajima--x
26 May 2026
Model Releases

Raon-Speech Technical Report

DGX agent

arXiv:2605.23912v1 Announce Type: cross Abstract: We present Raon-Speech, a top-performing 9B-parameter speech language model (SpeechLM) for English and Korean speech understanding, answering, and gen

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

RAW: Robust Avatar Watermarking -- Benchmarking and Baseline

DGX agent

arXiv:2605.23994v1 Announce Type: cross Abstract: Digital avatar watermarking presents unique challenges: avatars are routinely post-processed with background replacement, reframing, and format conver

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

READER: Reasoning-Enhanced AI-Generated Text Detection

DGX agent

arXiv:2605.25281v1 Announce Type: cross Abstract: Recent advances in large language models (LLMs) have made it increasingly difficult to distinguish human-written text from AI-generated content. Many

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

RealBench: Benchmarking Data-Driven Numerical Weather Forecasting Under Operational Conditions and Extreme Event Challenges

DGX agent

arXiv:2605.24945v1 Announce Type: cross Abstract: Accurate evaluation of weather forecasting models is critical for their reliable deployment in real-world applications. However, existing benchmarks p

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

RECTOR: Priority-Aware Rule-Based Reranking for Compliance-Aware Autonomous Driving Trajectory Selection

DGX agent

arXiv:2605.25095v1 Announce Type: new Abstract: Autonomous driving stacks must pick one trajectory from a multi-modal candidate set; choosing by model confidence ignores safety, traffic-law, and comfo

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

RED: Adaptive Real-Time DAG Scheduling for Robotic Inference under Environmental Dynamics

DGX agent

arXiv:2605.24044v1 Announce Type: new Abstract: Robots deployed in dynamic environments must contend with environment-driven changes that reshape computation at runtime: new tasks may appear, preceden

model-releasesarxiv-cs-ro
26 May 2026
Model Releases

Red-Teaming Claude Opus and ChatGPT-based Security Advisors for Trusted Execution Environments

DGX agent

arXiv:2602.19450v2 Announce Type: replace-cross Abstract: Trusted Execution Environments (TEEs) (e.g., Intel SGX and ArmTrustZone) aim to protect sensitive computation from a compromised operating sys

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Refining Context-Entangled Content Segmentation via Curriculum Selection and Anti-Curriculum Promotion

DGX agent

arXiv:2602.01183v2 Announce Type: replace-cross Abstract: Biological learning proceeds from easy to difficult tasks, gradually reinforcing perception and robustness. Inspired by this principle, we add

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

Reflect-Guard: Enhancing LLM Safeguards against Adversarial Prompts via Logical Self-Reflection

DGX agent

arXiv:2605.24834v1 Announce Type: cross Abstract: Large language model (LLM) safety classifiers such as Llama Guard are effective at detecting overtly harmful prompts but remain vulnerable to adversar

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Reinforcement Learning for Laser Additive Manufacturing Scan-Order Optimisation: A Bilevel Proxy--FEA Diagnostic Framework for Reward and World-Model Diagnosis

DGX agent

arXiv:2605.25063v1 Announce Type: new Abstract: Reinforcement learning offers a promising approach for scan-order optimisation in laser additive manufacturing, where sequential scan decisions critical

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

RepetitionCurse: Measuring and Understanding Router Imbalance in Mixture-of-Experts LLMs under DoS Stress

DGX agent

arXiv:2512.23995v2 Announce Type: replace-cross Abstract: Mixture-of-Experts architectures have become the standard for scaling large language models due to their superior parameter efficiency. To acc

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

RePlan-Bot: Multi-Level Replanning for Embodied Instruction Following

DGX agent

arXiv:2605.25851v1 Announce Type: new Abstract: Embodied instruction following (EIF) requires agents to understand and execute complex natural language commands within interactive 3D environments. Des

model-releasesarxiv-cs-ro
26 May 2026
Model Releases

Representation Without Control: Testing the Realization Effect in Language Models

DGX agent

arXiv:2605.25151v1 Announce Type: new Abstract: Large language models are increasingly used as behavioral simulators, but it remains unclear when their outputs reflect human-like cognitive mechanisms

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

RepSAM: Bridging Foundation Models to Robotic Vision via Representation-Guided Adaptation

DGX agent

arXiv:2605.25495v1 Announce Type: new Abstract: Robotic perception in unstructured environments remains challenging despite the zero-shot capabilities of foundation models such as SAM. This work attri

model-releasesarxiv-cs-ro
26 May 2026
Model Releases

Residual Drift Dominates Contradiction in Multi-Turn Constraint Reasoning

DGX agent

arXiv:2605.23940v1 Announce Type: new Abstract: How do multi-turn reasoning systems fail? The expected answer is logical contradiction, in which the system's maintained state becomes unsatisfiable. We

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Rethinking Continual Anomaly Detection on the Edge: Benchmarking Under Realistic Industrial Conditions

DGX agent

arXiv:2605.24251v1 Announce Type: new Abstract: Continual anomaly detection (CAD) addresses the need for industrial inspection systems to adapt to evolving production conditions, yet existing methods

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

Rethinking Weak Supervision in Anomaly Detection: A Comprehensive Benchmark

DGX agent

arXiv:2605.26068v1 Announce Type: cross Abstract: Weakly supervised anomaly detection (WSAD) has developed in three primary directions: incomplete, inexact, and inaccurate supervision. However, these

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Retrying vs Resampling in AI Control

DGX agent

arXiv:2605.26047v1 Announce Type: new Abstract: AI coding scaffolds like Claude Code and Codex use extit{retrying}: blocking actions flagged as risky and continuing the trajectory. We study retrying f

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Reward-free Alignment for Conflicting Objectives

DGX agent

arXiv:2602.02495v3 Announce Type: replace-cross Abstract: Direct alignment methods are increasingly used to align large language models (LLMs) with human preferences. However, many real-world alignmen

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Riemannian-Manifold Steering: Geometry-Aware Generative Autoencoders for Label-Free Steering

DGX agent

arXiv:2605.24942v1 Announce Type: cross Abstract: Steering a language model - intervening on its internal activations to change downstream behaviour - has recently expanded beyond linear interpolation

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

RL with Learnable Textual Feedback: A Bilevel Approach

DGX agent

arXiv:2605.24547v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards can improve LLM reasoning, but learning remains sample-inefficient when terminal rewards are sparse. This

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

RoboManipBaselines: A Unified Framework for Imitation Learning in Robotic Manipulation across Real and Simulation Environments

DGX agent

arXiv:2509.17057v3 Announce Type: replace Abstract: We present RoboManipBaselines, an open-source software framework for imitation learning research in robotic manipulation. The framework supports the

model-releasesarxiv-cs-ro
26 May 2026
Model Releases

Robust Fuzzy Multi-view Learning under View Conflict

DGX agent

arXiv:2605.24475v1 Announce Type: cross Abstract: Trusted multi-view classification aims to deliver reliable fusion for accurate predictions and has recently attracted substantial attention in both ac

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

RotMoLE: Enhancing Mixture of Low-Rank Experts through Rotational Gating Mechanism

DGX agent

arXiv:2605.25565v1 Announce Type: cross Abstract: While Large Language Models (LLMs) are commonly fine-tuned to handle domain-specific tasks before being applied to vertical applications, adapting the

model-releasesarxiv-cs-cl
26 May 2026
Model Releases

SafeCtrl-RL: Inference-Time Adaptive Behaviour Control for LLM Dialogue via RL-Driven Prompt Optimisation

DGX agent

arXiv:2605.25984v1 Announce Type: cross Abstract: Ensuring safe and contextually appropriate behaviour in Large Language Models (LLMs) remains a critical challenge for real-world deployment. We presen

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Safety Generalization Under Distribution Shift in Safe Reinforcement Learning: A Diabetes Testbed

DGX agent

arXiv:2601.21094v2 Announce Type: replace-cross Abstract: Safe Reinforcement Learning (RL) algorithms are typically evaluated under fixed training conditions. We investigate whether training-time safe

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

SafetyRepro: Configuration-Conditional Rank Instability on Alignment Benchmarks

DGX agent

arXiv:2605.25492v1 Announce Type: new Abstract: Pairwise model comparisons drawn from foundation-model benchmarks ('A is safer than B') are read as quantitative verdicts but hinge on harness choices b

model-releasesarxiv-cs-lg
26 May 2026
← Previous
1…266267268269270…472
Next →