AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent
84,619Total entries
1Added by human
84,618Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,603 results
26 May 2026

PACZero: PAC-Private Fine-Tuning of Language Models via Sign Quantization

Model ReleasesDGX agent

arXiv:2605.06505v2 Announce Type: replace-cross Abstract: We introduce PACZero, a family of PAC-private zeroth-order mechanisms for fine-tuning large language models that delivers usable utility at I(

Palette: A Modular, Controllable, and Efficient Framework for On-demand Authorized Safety Alignment Relaxation in LLMs

Model ReleasesDGX agent

arXiv:2605.24154v1 Announce Type: new Abstract: Current safety alignment of foundation models largely follows a one-size-fits-all paradigm, applying the same refusal policy across users and contexts.

PALoRA: Projection-Adaptive LoRA for Preserving Reasoning in Large Language Models

Model Releases

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2605.24549v1 Announce Type: new Abstract: Efficiently updating Large Language Models (LLMs) with new or evolving factual knowledge remains a central challenge, as even parameter-efficient adapta

Parameter-Efficient CT Reconstruction via Deep Graph Laplacian Regularization

Model ReleasesDGX agent

arXiv:2605.25348v1 Announce Type: cross Abstract: Low-dose computed tomography (LDCT) reconstruction faces a critical tradeoff between reconstruction quality and resource requirements. While recent de

Parameter Efficient Multi-Class Intelligent Scheduling for Multimodal Online Distributed Industrial Anomaly Detection

Model ReleasesDGX agent

arXiv:2605.23984v1 Announce Type: cross Abstract: Industrial anomaly detection has attracted significant attention as a fundamental challenge in industrial systems. The rapid advancement of heterogene

Parameter-Efficient VLMs for Gastrointestinal Endoscopy: Medical Image Generation and Clinical Visual Question Answering

Model ReleasesDGX agent

arXiv:2605.24792v1 Announce Type: cross Abstract: The major limitations of gastrointestinal (GI) endoscopy AI systems arise from a shortage of annotated data, strict privacy policies, and significant

ParkourFormer: Integrating Predictive Supervision and Sequence Modeling into Parkour Locomotion

Model ReleasesDGX agent

arXiv:2605.25782v1 Announce Type: new Abstract: Humanoid parkour requires locomotion policies to coordinate whole-body dynamics across rapidly changing terrains such as stairs, gaps, slopes, and obsta

Partition of Unity Neural Networks for Interpretable Classification with Explicit Class Regions

Model ReleasesDGX agent

arXiv:2602.00511v2 Announce Type: replace Abstract: Despite their empirical success, neural network classifiers remain difficult to interpret. In softmax-based models, class regions are defined implic

Partner-Aware Hierarchical Skill Discovery for Robust Human-AI Collaboration

Model ReleasesDGX agent

arXiv:2605.24352v1 Announce Type: new Abstract: Multi-agent collaboration, especially in human-AI teaming, requires agents that can adapt to novel partners with diverse and dynamic behaviors. Conventi

PDEInvBench: A Comprehensive Dataset and Design Space Exploration of Neural Networks for PDE Inverse Problems

Model ReleasesDGX agent

arXiv:2605.25353v1 Announce Type: new Abstract: Inverse problems in partial differential equations (PDEs) involve estimating the physical parameters of a system from observed spatiotemporal solution f

PEDESTRIANQA: A Benchmark for Vision-Language Models on Pedestrian Intention and Trajectory Prediction

Model ReleasesDGX agent

arXiv:2605.24562v1 Announce Type: cross Abstract: Pedestrian intention and trajectory prediction are critical for the safe deployment of autonomous driving systems, directly influencing navigation dec

PennySynth: RAG-Driven Data Synthesis for Automated Quantum Code Generation

Model ReleasesDGX agent

arXiv:2605.25572v1 Announce Type: cross Abstract: The growing complexity of quantum programming frameworks has exposed a critical limitation in existing large language model (LLM)-based code assistant

Personalize-then-Store: Benchmarking and Learning Personalized Memory for Long-horizon Agents

Model ReleasesDGX agent

arXiv:2605.25535v1 Announce Type: new Abstract: Existing large language model (LLM) based memory systems apply universal, static policies that overlook a fundamental reality: the contexts that are wor

Personalized Federated Learning by Energy-Efficient UAV Communications

Model ReleasesDGX agent

arXiv:2605.25212v1 Announce Type: new Abstract: Federated learning (FL) is an effective paradigm for enhancing the learning capability of edge devices while preserving data privacy. In geographically

Physen-Noise2Noise: Physics-Guided Self-Supervised Defocus Deblurring with Bias Correction under Low-Light Conditions

Model ReleasesDGX agent

arXiv:2605.24590v1 Announce Type: cross Abstract: Low-light, long-exposure defocus deblurring remains a challenging problem due to the simultaneous presence of severe blur and complex biased noise. Ex

PiXTime: A Model for Federated Time Series Forecasting with Heterogeneous Data across Nodes

Model ReleasesDGX agent

arXiv:2601.05613v2 Announce Type: replace-cross Abstract: While collaborative forecasting on distributed time series is highly desirable, directly pooling localized datasets is often impractical due t

Polymorphism Is Rotation: Operational Mechanistic Interpretability from a Two-Layer Transformer to Pythia-70m

Model ReleasesDGX agent

arXiv:2605.24577v1 Announce Type: cross Abstract: Independently trained transformers compute the same function in residual-stream bases that differ by a uniform random rotation on SO(d_{model}). We ca

PolySAE: Modeling Feature Interactions in Sparse Autoencoders via Polynomial Decoding

Model ReleasesDGX agent

arXiv:2602.01322v2 Announce Type: replace-cross Abstract: Sparse autoencoders (SAEs) interpret neural network representations by decomposing activations into sparse combinations of dictionary atoms. H

Position: AI for Science Should Treat Measurement-to-Dataset Pipelines as Inference Components

Model ReleasesDGX agent

arXiv:2605.24558v1 Announce Type: new Abstract: AI for Science (AI4Science) workflows often treat the released dataset as a fixed interface to the underlying system. However, in domains relying on ind

Pragmatic Reasoning improves LLM Code Generation

Model ReleasesDGX agent

arXiv:2502.15835v5 Announce Type: replace-cross Abstract: Pragmatic reasoning helps interlocutors infer intended meaning from ambiguous or underspecified messages by considering shared context and cou

Quantifying the Impact of Translation Errors on Multilingual LLM Evaluation

Model ReleasesDGX agent

arXiv:2605.24904v1 Announce Type: new Abstract: Machine-translated benchmarks are widely used to assess the multilingual capabilities of large language models (LLMs), yet translation errors in these b

Quaternion Self-Attention with Shared Scores

Model ReleasesDGX agent

arXiv:2605.24920v1 Announce Type: cross Abstract: Quaternion neural networks are parameter-efficient and model multidimensional dependencies by representing four related features as a single entity. H

Querying structural and functional niches on spatial transcriptomics data

Model ReleasesDGX agent

arXiv:2410.10652v4 Announce Type: replace-cross Abstract: Cells in multicellular organisms coordinate to form structural and functional niches. With spatial transcriptomics (ST) enabling gene expressi

QUEST: Training Frontier Deep Research Agents with Fully Synthetic Tasks

Model ReleasesDGX agent

arXiv:2605.24218v1 Announce Type: new Abstract: Deep research agents extend the role of search engines from retrieving keyword-matched pages to synthesizing knowledge, fundamentally changing how human

QUIET: A Multi-Blank Cascaded Story Cloze Benchmark for LLM Creative Generation Capability

Model ReleasesDGX agent

arXiv:2605.25955v1 Announce Type: cross Abstract: Large language models (LLMs) face a dual challenge in creative capability evaluation: existing benchmarks (e.g., Story Cloze Test, HellaSwag) measure

Qwen 3.7 Max is now supported in Hermes Agent

Model ReleasesDGX agent

Nous Research has added support for Qwen 3.7 Max, a large language model, within their Hermes Agent framework. This integration enables users to leverage Qwen 3.7 Max's capabilities when building or d

Qwen3.7 Max now available in Go - text only - 1M context - smartest model in the Qwen family to date

Model ReleasesDGX agent

Alibaba's Qwen team has released Qwen3.7 Max, their most advanced model to date, now available for use with Go programming language support. The model features a 1 million token context window and is

ran my first benchmark this weekend (longmemeval) mostly to test activegraph, learned a lot! - this is a stepping stone to show the event ba…

Model ReleasesDGX agent

ran my first benchmark this weekend (longmemeval) mostly to test activegraph, learned a lot! - this is a stepping stone to show the event based agent system works. the AI convinced me not to start wit

Raon-Speech Technical Report

Model ReleasesDGX agent

arXiv:2605.23912v1 Announce Type: cross Abstract: We present Raon-Speech, a top-performing 9B-parameter speech language model (SpeechLM) for English and Korean speech understanding, answering, and gen

RAW: Robust Avatar Watermarking -- Benchmarking and Baseline

Model ReleasesDGX agent

arXiv:2605.23994v1 Announce Type: cross Abstract: Digital avatar watermarking presents unique challenges: avatars are routinely post-processed with background replacement, reframing, and format conver

READER: Reasoning-Enhanced AI-Generated Text Detection

Model ReleasesDGX agent

arXiv:2605.25281v1 Announce Type: cross Abstract: Recent advances in large language models (LLMs) have made it increasingly difficult to distinguish human-written text from AI-generated content. Many

RealBench: Benchmarking Data-Driven Numerical Weather Forecasting Under Operational Conditions and Extreme Event Challenges

Model ReleasesDGX agent

arXiv:2605.24945v1 Announce Type: cross Abstract: Accurate evaluation of weather forecasting models is critical for their reliable deployment in real-world applications. However, existing benchmarks p

RECTOR: Priority-Aware Rule-Based Reranking for Compliance-Aware Autonomous Driving Trajectory Selection

Model ReleasesDGX agent

arXiv:2605.25095v1 Announce Type: new Abstract: Autonomous driving stacks must pick one trajectory from a multi-modal candidate set; choosing by model confidence ignores safety, traffic-law, and comfo

RED: Adaptive Real-Time DAG Scheduling for Robotic Inference under Environmental Dynamics

Model ReleasesDGX agent

arXiv:2605.24044v1 Announce Type: new Abstract: Robots deployed in dynamic environments must contend with environment-driven changes that reshape computation at runtime: new tasks may appear, preceden

Red-Teaming Claude Opus and ChatGPT-based Security Advisors for Trusted Execution Environments

Model ReleasesDGX agent

arXiv:2602.19450v2 Announce Type: replace-cross Abstract: Trusted Execution Environments (TEEs) (e.g., Intel SGX and ArmTrustZone) aim to protect sensitive computation from a compromised operating sys

Refining Context-Entangled Content Segmentation via Curriculum Selection and Anti-Curriculum Promotion

Model ReleasesDGX agent

arXiv:2602.01183v2 Announce Type: replace-cross Abstract: Biological learning proceeds from easy to difficult tasks, gradually reinforcing perception and robustness. Inspired by this principle, we add

Reflect-Guard: Enhancing LLM Safeguards against Adversarial Prompts via Logical Self-Reflection

Model ReleasesDGX agent

arXiv:2605.24834v1 Announce Type: cross Abstract: Large language model (LLM) safety classifiers such as Llama Guard are effective at detecting overtly harmful prompts but remain vulnerable to adversar

Reinforcement Learning for Laser Additive Manufacturing Scan-Order Optimisation: A Bilevel Proxy--FEA Diagnostic Framework for Reward and World-Model Diagnosis

Model ReleasesDGX agent

arXiv:2605.25063v1 Announce Type: new Abstract: Reinforcement learning offers a promising approach for scan-order optimisation in laser additive manufacturing, where sequential scan decisions critical

RepetitionCurse: Measuring and Understanding Router Imbalance in Mixture-of-Experts LLMs under DoS Stress

Model ReleasesDGX agent

arXiv:2512.23995v2 Announce Type: replace-cross Abstract: Mixture-of-Experts architectures have become the standard for scaling large language models due to their superior parameter efficiency. To acc

RePlan-Bot: Multi-Level Replanning for Embodied Instruction Following

Model ReleasesDGX agent

arXiv:2605.25851v1 Announce Type: new Abstract: Embodied instruction following (EIF) requires agents to understand and execute complex natural language commands within interactive 3D environments. Des

Representation Without Control: Testing the Realization Effect in Language Models

Model ReleasesDGX agent

arXiv:2605.25151v1 Announce Type: new Abstract: Large language models are increasingly used as behavioral simulators, but it remains unclear when their outputs reflect human-like cognitive mechanisms

RepSAM: Bridging Foundation Models to Robotic Vision via Representation-Guided Adaptation

Model ReleasesDGX agent

arXiv:2605.25495v1 Announce Type: new Abstract: Robotic perception in unstructured environments remains challenging despite the zero-shot capabilities of foundation models such as SAM. This work attri

Residual Drift Dominates Contradiction in Multi-Turn Constraint Reasoning

Model ReleasesDGX agent

arXiv:2605.23940v1 Announce Type: new Abstract: How do multi-turn reasoning systems fail? The expected answer is logical contradiction, in which the system's maintained state becomes unsatisfiable. We

Rethinking Continual Anomaly Detection on the Edge: Benchmarking Under Realistic Industrial Conditions

Model ReleasesDGX agent

arXiv:2605.24251v1 Announce Type: new Abstract: Continual anomaly detection (CAD) addresses the need for industrial inspection systems to adapt to evolving production conditions, yet existing methods

Rethinking Weak Supervision in Anomaly Detection: A Comprehensive Benchmark

Model ReleasesDGX agent

arXiv:2605.26068v1 Announce Type: cross Abstract: Weakly supervised anomaly detection (WSAD) has developed in three primary directions: incomplete, inexact, and inaccurate supervision. However, these

Retrying vs Resampling in AI Control

Model ReleasesDGX agent

arXiv:2605.26047v1 Announce Type: new Abstract: AI coding scaffolds like Claude Code and Codex use extit{retrying}: blocking actions flagged as risky and continuing the trajectory. We study retrying f

Reward-free Alignment for Conflicting Objectives

Model ReleasesDGX agent

arXiv:2602.02495v3 Announce Type: replace-cross Abstract: Direct alignment methods are increasingly used to align large language models (LLMs) with human preferences. However, many real-world alignmen

Riemannian-Manifold Steering: Geometry-Aware Generative Autoencoders for Label-Free Steering

Model ReleasesDGX agent

arXiv:2605.24942v1 Announce Type: cross Abstract: Steering a language model - intervening on its internal activations to change downstream behaviour - has recently expanded beyond linear interpolation

RL with Learnable Textual Feedback: A Bilevel Approach

Model ReleasesDGX agent

arXiv:2605.24547v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards can improve LLM reasoning, but learning remains sample-inefficient when terminal rewards are sparse. This

RoboManipBaselines: A Unified Framework for Imitation Learning in Robotic Manipulation across Real and Simulation Environments

Model ReleasesDGX agent

arXiv:2509.17057v3 Announce Type: replace Abstract: We present RoboManipBaselines, an open-source software framework for imitation learning research in robotic manipulation. The framework supports the

Robust Fuzzy Multi-view Learning under View Conflict

Model ReleasesDGX agent

arXiv:2605.24475v1 Announce Type: cross Abstract: Trusted multi-view classification aims to deliver reliable fusion for accurate predictions and has recently attracted substantial attention in both ac

RotMoLE: Enhancing Mixture of Low-Rank Experts through Rotational Gating Mechanism

Model ReleasesDGX agent

arXiv:2605.25565v1 Announce Type: cross Abstract: While Large Language Models (LLMs) are commonly fine-tuned to handle domain-specific tasks before being applied to vertical applications, adapting the

SafeCtrl-RL: Inference-Time Adaptive Behaviour Control for LLM Dialogue via RL-Driven Prompt Optimisation

Model ReleasesDGX agent

arXiv:2605.25984v1 Announce Type: cross Abstract: Ensuring safe and contextually appropriate behaviour in Large Language Models (LLMs) remains a critical challenge for real-world deployment. We presen

Safety Generalization Under Distribution Shift in Safe Reinforcement Learning: A Diabetes Testbed

Model ReleasesDGX agent

arXiv:2601.21094v2 Announce Type: replace-cross Abstract: Safe Reinforcement Learning (RL) algorithms are typically evaluated under fixed training conditions. We investigate whether training-time safe

SafetyRepro: Configuration-Conditional Rank Instability on Alignment Benchmarks

Model ReleasesDGX agent

arXiv:2605.25492v1 Announce Type: new Abstract: Pairwise model comparisons drawn from foundation-model benchmarks ('A is safer than B') are read as quantitative verdicts but hinge on harness choices b

ScaleAcross Explorer: Exploring Communication Optimization for Scale-Across AI Model Training

Model ReleasesDGX agent

arXiv:2605.24326v1 Announce Type: cross Abstract: The rapid scaling of large language model training requires distributing GPU resources across multiple data center buildings and regions. We refer to

Schema-Grounded LLM Extraction for FHIR Patient Digital Twins

Model ReleasesDGX agent

arXiv:2601.05847v2 Announce Type: replace Abstract: We revisit the problem of constructing interoperable patient digital twins from unstructured electronic health records (EHRs) and argue that the tas

Second Guess: Detecting Uncertainty Through Abstention and Answer Stability in Small Language Models

Model ReleasesDGX agent

arXiv:2605.25394v1 Announce Type: new Abstract: Large language models often generate confident but incorrect answers rather than abstaining when uncertain. This problem is particularly acute for small

Security in the Fine-Tuning Lifecycle of Large Language Models: Threats, Defenses,Evaluation, and Future Directions

Model ReleasesDGX agent

arXiv:2605.25073v1 Announce Type: cross Abstract: Background: Fine-tuning is central to adapting pre-trained Large Language Models (LLMs) to downstream tasks, but its reliance on training data, parame

SemanticZip: A Pilot Framework for Lossy Text Compression with LLMs as Semantic Decompressors

Model ReleasesDGX agent

arXiv:2605.24541v1 Announce Type: cross Abstract: Text compression for large language model (LLM) systems is usually framed as token deletion, retrieval, summarization, or exact reconstruction. We stu

← Previous
1…213214215216217…377
Next →