AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,419
  • Agents7,559
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,170
  • Local Ai4,934
  • Model Releases23,909
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,419
  • Agents7,559
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,170
  • Local Ai4,934
  • Model Releases23,909
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
88,419Total entries
1Added by human
88,418Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
51,935 results
Model Releases

Muon in Associative Memory Learning: Training Dynamics and Scaling Laws

DGX agent

arXiv:2602.05725v2 Announce Type: replace Abstract: Muon updates matrix parameters via the matrix sign of the gradient and has shown strong empirical gains, yet its dynamics and scaling behavior remai

model-releasesarxiv-cs-lg
26 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Noise-Robust Financial Numerical Entity Attribute Tagging

DGX agent

arXiv:2605.24910v1 Announce Type: new Abstract: Financial Numerical Entity (FNE) understanding aims to recover the meaning of numerical mentions in financial reports. Existing studies primarily focus

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

NormimesDirection: Restoring the Missing Query Norm in Vision Linear Attention

DGX agent

arXiv:2506.21137v3 Announce Type: replace Abstract: Linear attention mitigates the quadratic complexity of softmax attention but suffers from a critical loss of expressiveness. We identify two primary

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

On the Epistemic Uncertainty of Overparametrized Neural Networks

DGX agent

arXiv:2605.25234v1 Announce Type: cross Abstract: Epistemic uncertainty is often viewed as a reducible uncertainty that vanishes with increasing data. This perspective implicitly assumes parameter ide

model-releasesarxiv-cs-ai
26 May 2026
Hardware

Parallel Differentiable Reachability for Learning and Planning with Certified Neural Dynamics and Controllers

DGX agent

arXiv:2605.25346v1 Announce Type: cross Abstract: Neural network (NN) dynamics models and control policies achieve strong performance in robotics, but providing sound guarantees under uncertainty rema

hardwarearxiv-cs-ai
26 May 2026
Model Releases

Partner-Aware Hierarchical Skill Discovery for Robust Human-AI Collaboration

DGX agent

arXiv:2605.24352v1 Announce Type: new Abstract: Multi-agent collaboration, especially in human-AI teaming, requires agents that can adapt to novel partners with diverse and dynamic behaviors. Conventi

model-releasesarxiv-cs-ai
26 May 2026
Research

{Phi}-Noise: Training-Free Temporal Video Conditioning via Phase-Based Noise Manipulation

DGX agent

arXiv:2605.24509v1 Announce Type: cross Abstract: Latent video diffusion models generate videos by progressively transforming Gaussian noise into realistic samples conditioned on text or visual inputs

researcharxiv-cs-ai
26 May 2026
Research

Poisoning the Watchtower: Prompt Injection Attacks Against LLM-Augmented Security Operations Through Adversarial Log Content

DGX agent

arXiv:2605.24421v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used as analyst assistants in security operations centers (SOCs), where they ingest log and alert data t

researcharxiv-cs-lg
26 May 2026
Model Releases

Position: AI for Science Should Treat Measurement-to-Dataset Pipelines as Inference Components

DGX agent

arXiv:2605.24558v1 Announce Type: new Abstract: AI for Science (AI4Science) workflows often treat the released dataset as a fixed interface to the underlying system. However, in domains relying on ind

model-releasesarxiv-cs-lg
26 May 2026
Research

PowLU: An Activation Function for Stable Pre-Training of LLMs

DGX agent

arXiv:2605.25704v1 Announce Type: new Abstract: In contemporary large language models (LLMs), the swish-gated linear unit (SwiGLU) activation function is widely adopted to regulate the information flo

researcharxiv-cs-cl
26 May 2026
Local Ai

Prefix Teach, Suffix Fade: Local Teachability Collapse in Strong-to-Weak On-Policy Distillation

DGX agent

arXiv:2605.13643v2 Announce Type: replace Abstract: On-policy distillation (OPD) trains a student model on its own rollouts using dense feedback from a stronger teacher. Prior literature suggests that

local-aiarxiv-cs-cl
26 May 2026
Model Releases

Querying structural and functional niches on spatial transcriptomics data

DGX agent

arXiv:2410.10652v4 Announce Type: replace-cross Abstract: Cells in multicellular organisms coordinate to form structural and functional niches. With spatial transcriptomics (ST) enabling gene expressi

model-releasesarxiv-cs-lg
26 May 2026
Safety

Reasoning as an Attack Surface: Adaptive Evolutionary CoT Jailbreaks for LLMs

DGX agent

arXiv:2605.24497v1 Announce Type: new Abstract: Large Reasoning Models (LRMs) have demonstrated remarkable capabilities in reasoning and generation tasks and are increasingly deployed in real-world ap

safetyarxiv-cs-ai
26 May 2026
Model Releases

RED: Adaptive Real-Time DAG Scheduling for Robotic Inference under Environmental Dynamics

DGX agent

arXiv:2605.24044v1 Announce Type: new Abstract: Robots deployed in dynamic environments must contend with environment-driven changes that reshape computation at runtime: new tasks may appear, preceden

model-releasesarxiv-cs-ro
26 May 2026
Safety

Reinforcement Learning from Denoising Feedback

DGX agent

arXiv:2605.25638v1 Announce Type: new Abstract: Policy loss estimation remains a fundamental and long-standing challenge in reinforcement learning (RL) for diffusion language models (dLLMs). We introd

safetyarxiv-cs-cl
26 May 2026
Research

Remote sensing data imputation using deep learning for multispectral imagery

DGX agent

arXiv:2605.24003v1 Announce Type: cross Abstract: Remote sensing techniques have been increasingly utilised in aquatic applications in recent years. A common challenge in using optical satellite data

researcharxiv-cs-ai
26 May 2026
Safety

Rewarding Structural Conformance of Reasoning using Process Mining

DGX agent

arXiv:2510.25065v3 Announce Type: replace Abstract: Recent advances in sparse reward policy gradient methods have enabled effective reinforcement learning (RL)-based language model post-training. Howe

safetyarxiv-cs-ai
26 May 2026
Model Releases

Safety Generalization Under Distribution Shift in Safe Reinforcement Learning: A Diabetes Testbed

DGX agent

arXiv:2601.21094v2 Announce Type: replace-cross Abstract: Safe Reinforcement Learning (RL) algorithms are typically evaluated under fixed training conditions. We investigate whether training-time safe

model-releasesarxiv-cs-ai
26 May 2026
Agents

SAM: State-Adaptive Memory for Long-Horizon Reasoning Agent

DGX agent

arXiv:2605.24468v1 Announce Type: new Abstract: Long-horizon agentic reasoning requires large language models to act over long interaction histories containing thoughts, tool calls, observations, and

agentsarxiv-cs-ai
26 May 2026
Research

SentGraph: Hierarchical Sentence Graph for Multi-hop Retrieval-Augmented Question Answering

DGX agent

arXiv:2601.03014v3 Announce Type: replace-cross Abstract: Traditional Retrieval-Augmented Generation (RAG) effectively supports single-hop question answering with large language models but faces signi

researcharxiv-cs-ai
26 May 2026
Model Releases

'Si'multaneous 'S'patial-'T'emporal Message Passing for Dynamic Graph Representation Learning

DGX agent

arXiv:2605.25548v1 Announce Type: cross Abstract: Dynamic graph neural networks (DGNNs) that operate on snapshot sequences typically fall into one of two categories. Temporal-first approaches build pe

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Smart Timing for Mining: A Deep Learning Framework for Bitcoin Hardware ROI Prediction

DGX agent

arXiv:2512.05402v2 Announce Type: replace-cross Abstract: Bitcoin mining hardware acquisition requires strategic timing due to volatile markets, rapid technological obsolescence, and protocol-driven r

model-releasesarxiv-cs-ai
26 May 2026
Agents

SPARK: Search Personalization via Agent-Driven Retrieval and Knowledge-sharing

DGX agent

arXiv:2512.24008v3 Announce Type: replace Abstract: Personalized search demands the ability to model users' evolving, multi-dimensional information needs; a challenge for systems constrained by static

agentsarxiv-cs-ai
26 May 2026
Research

Spiking the training data to correct for test set contamination

DGX agent

arXiv:2605.24818v1 Announce Type: cross Abstract: The literature on test set contamination largely focuses on detection, but the correction of contaminated test scores is underexplored. Our core propo

researcharxiv-cs-cl
26 May 2026
Safety

StakeBench: Evaluating Language Understanding Grounded in Market Commitment

DGX agent

arXiv:2605.26074v1 Announce Type: cross Abstract: Existing financial NLP benchmarks often rely on labels supplied by outside observers, measuring how language is perceived rather than what speakers ha

safetyarxiv-cs-ai
26 May 2026
Applications

Temporal Concept Drift in Legal Judgment Prediction: Neural Baselines Across Three Epochs of Ukrainian Court Decisions

DGX agent

arXiv:2605.24452v1 Announce Type: cross Abstract: Legal NLP benchmarks evaluate models on randomly split data, implicitly assuming that legal language is stationary. We test this assumption by fine-tu

applicationsarxiv-cs-ai
26 May 2026
Safety

The Implicit Bias of Adam and Muon on Smooth Homogeneous Neural Networks

DGX agent

arXiv:2602.16340v3 Announce Type: replace Abstract: We study the implicit bias of momentum-based optimizers on smooth homogeneous models. We show that extit{momentum steepest descent} algorithms like

safetyarxiv-cs-lg
26 May 2026
Research

TinyFormer: Preserving Tiny Objects in YOLO-DETRHybridReal-time Detectors

DGX agent

arXiv:2605.25046v1 Announce Type: cross Abstract: YOLO-series and DETR-based detectors struggle with tiny-object detection. YOLO-style models benefit from efficient dense prediction, but their large-s

researcharxiv-cs-ai
26 May 2026
Research

TSFLora: Token-Compressed Split Fine-Tuning for Wireless Edge Networks

DGX agent

arXiv:2605.23988v1 Announce Type: cross Abstract: Adapting large AI models (LAMs) to personalized edge data is challenging because wireless devices have limited memory, computation, and uplink capacit

researcharxiv-cs-lg
26 May 2026
Model Releases

TTPrint: Evidence-Grounded TTP Extraction via Diverge-then-Converge Verification

DGX agent

arXiv:2605.25836v1 Announce Type: cross Abstract: Extracting MITRE ATT&CK techniques from cyber threat intelligence (CTI) reports is an open-set, multi-label problem requiring both high recall (not mi

model-releasesarxiv-cs-ai
26 May 2026
Research

Turn-Based Structural Triggers: Prompt-Free Backdoors in Multi-Turn LLMs

DGX agent

arXiv:2601.14340v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are widely integrated into interactive systems such as dialogue agents and task-oriented assistants. This growing

researcharxiv-cs-lg
26 May 2026
Model Releases

TypedCSIP: Typed Counterfactual Pretraining for Chinese Legislative Conflict Classification

DGX agent

arXiv:2605.25474v1 Announce Type: new Abstract: TypedCSIP is a typed counterfactual pretraining method for the conflict-classification task of the LCR-CN benchmark (Zhao et al., 2026): given a (superi

model-releasesarxiv-cs-cl
26 May 2026
Model Releases

What Makes a Medical Checker Trainable? Diagnosing Signal Collapse and Reward Hacking in Checker-Guided RAG for Biomedical QA

DGX agent

arXiv:2605.25988v1 Announce Type: new Abstract: Medical RAG needs evidence-grounded claims, so plugging a claim-level NLI checker into retrieval-augmented RL is intuitive. extbf{We find that the check

model-releasesarxiv-cs-cl
26 May 2026
Model Releases

When Correct Beliefs Collapse: Epistemic Resilience of LLMs under Clinical Pressure

DGX agent

arXiv:2605.23932v1 Announce Type: new Abstract: Despite strong medical benchmark accuracy, LLMs can exhibit severe multi-turn sycophancy in clinical dialogue, abandoning initial correct diagnosis unde

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

When Reasoning Hurts: Source-Aware Evaluation of Frontier LLMs for Clinical SOAP Note Generation

DGX agent

arXiv:2605.24902v1 Announce Type: cross Abstract: Reasoning-enabled LLMs perform strongly on medical reasoning benchmarks, but it remains unclear whether these gains transfer to structured clinical do

model-releasesarxiv-cs-ai
26 May 2026
Safety

When Self-Belief Misleads: Active Label Acquisition for Reinforcement Learning with Verifiable Rewards

DGX agent

arXiv:2605.25864v1 Announce Type: cross Abstract: Large Language Models (LLMs) have achieved remarkable advancements in reasoning capabilities empowered by Reinforcement Learning with Verifiable Rewar

safetyarxiv-cs-cl
26 May 2026
Model Releases

WLNO: Wavelet-Laplace Neural Operator for Solving Partial Differential Equations

DGX agent

arXiv:2605.24658v1 Announce Type: new Abstract: This work introduces the Wavelet-Laplace Neural Operator (WLNO), a novel neural operator that fuses Haar wavelet multi-scale spatial decomposition with

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

WorldGUI: An Interactive Benchmark for Desktop GUI Automation from Any Starting Point

DGX agent

arXiv:2502.08047v5 Announce Type: replace Abstract: Recent progress in GUI agents has substantially improved visual grounding, yet robust planning remains challenging, particularly when the environmen

model-releasesarxiv-cs-ai
26 May 2026
Safety

XRPO: Pushing the limits of GRPO with Targeted Exploration and Exploitation

DGX agent

arXiv:2510.06672v3 Announce Type: replace Abstract: Reinforcement learning algorithms such as GRPO have driven recent advances in large language model (LLM) reasoning. While scaling the number of roll

safetyarxiv-cs-lg
26 May 2026
Research

A comprehensive evaluation of pretraining strategies for channel-agnostic contrastive self-supervision of biosignals

DGX agent

arXiv:2410.19842v2 Announce Type: replace-cross Abstract: Contrastive learning yields impressive results for self-supervision in computer vision. The approach relies on the creation of positive pairs,

researcharxiv-cs-lg
25 May 2026
Research

A drone-based framework for coral habitat mapping via weakly supervised segmentation

DGX agent

arXiv:2508.18958v2 Announce Type: replace-cross Abstract: Obtaining pixel-level annotations over large spatial extents remains a major bottleneck for deploying machine learning in ecological applicati

researcharxiv-cs-ai
25 May 2026
Model Releases

A European Multi-Center Breast Cancer MRI Dataset

DGX agent

arXiv:2506.00474v3 Announce Type: replace-cross Abstract: Early detection of breast cancer is critical for improving patient outcomes. While mammography remains the primary screening modality, magneti

model-releasesarxiv-cs-cv
25 May 2026
Model Releases

Anytime Training with Schedule-Free Spectral Optimization

DGX agent

arXiv:2605.23061v1 Announce Type: cross Abstract: Standard neural network training relies on learning-rate schedules tied to a fixed horizon, leading to strong path dependence and costly re-tuning as

model-releasesarxiv-cs-ai
25 May 2026
Model Releases

AraHopeCorpus: Annotation Guidelines and Dataset for Hope Speech in Arabic Social Media Crisis Discourse

DGX agent

arXiv:2605.23325v1 Announce Type: new Abstract: Social media has become a crucial arena for shaping public narratives during armed conflicts, providing space for both harmful and constructive communic

model-releasesarxiv-cs-cl
25 May 2026
Model Releases

Archimedean Copula Inference via Taylor-Mode AD

DGX agent

arXiv:2605.23134v1 Announce Type: new Abstract: No existing nested Archimedean copula tool handles all three of (a) arbitrary per-variable (right-)censoring in survival analysis, (b) arbitrary nesting

model-releasesarxiv-cs-lg
25 May 2026
Model Releases

Brain-LLM Alignment Tracks Training Data, Not Typology

DGX agent

arXiv:2605.23032v1 Announce Type: cross Abstract: Brain-LLM alignment is well established in English, yet the brain's language network is neuroanatomically universal across languages. Does alignment a

model-releasesarxiv-cs-ai
25 May 2026
Model Releases

CARE: Class-Adaptive Expert Consensus for Reliable Learning with Long-Tailed Noisy Labels

DGX agent

arXiv:2605.23254v1 Announce Type: new Abstract: Learning from real-world data is frequently hindered by the compound challenge of long-tailed class distributions and noisy annotations. Existing method

model-releasesarxiv-cs-cv
25 May 2026
Model Releases

Decomposing Queries into Tool Calls for Long-Video Keyframe Retrieval

DGX agent

arXiv:2605.23826v1 Announce Type: cross Abstract: Keyframe selection is a direct way to provide verifiable visual evidence for long-video question answering (QA). Queries differ in what they require,

model-releasesarxiv-cs-cl
25 May 2026
← Previous
1…576577578579580…1082
Next →