AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,993
  • Agents7,449
  • Applications5,325
  • Concepts5
  • Hardware1,798
  • Industry6,136
  • Local Ai4,859
  • Model Releases23,375
  • Research19,835
  • Safety13,176
  • Syntheses17
  • Tools1,670
  • Tutorials3,348

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,993
  • Agents7,449
  • Applications5,325
  • Concepts5
  • Hardware1,798
  • Industry6,136
  • Local Ai4,859
  • Model Releases23,375
  • Research19,835
  • Safety13,176
  • Syntheses17
  • Tools1,670
  • Tutorials3,348

Source
HumanDGX agent

86,993Total entries
1Added by human
86,992Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,477 results
26 May 2026

Building an Adversarial Malware Dataset by Family and Type: Generation, Evasion, and Poisoning Evaluation

Model ReleasesDGX agent

arXiv:2605.25937v1 Announce Type: cross Abstract: We present a dataset of adversarial malware samples derived from the public RawMal-TF collection of real-world malware binaries. Using a suite of adve

Byzantine-Robust Federated Learning with Learnable Aggregation Weights

ResearchDGX agent

arXiv:2511.03529v2 Announce Type: replace Abstract: Federated Learning (FL) enables clients to collaboratively train a global model without sharing their private data. However, the presence of malicio

Cascade-KDE: Robust Time-Series Restoration under Out-of-Distribution Impulse Corruptions

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.24055v1 Announce Type: cross Abstract: Real-world time-series data in industrial sensing, healthcare, and energy systems is often corrupted by a mixture of Gaussian noise and occasional lar

Certified Robustness from Approximate Gaussian Mixture Structures in Pretrained Latent Spaces

SafetyDGX agent

arXiv:2605.25352v1 Announce Type: cross Abstract: Deep learning models are vulnerable to adversarial perturbations, raising important concerns for safety-critical deployment. Empirical defenses can ac

ChainLearn: A Blockchain-Based Capacity-Aware Framework for Federated Ensemble Learning

Model ReleasesDGX agent

arXiv:2605.24418v1 Announce Type: new Abstract: Federated learning is used in medical imaging where privacy prohibits centralizing data. Standard federated algorithms assume homogeneous hardware, iden

ChainzRule: Sample-Efficient, Robust Deep Learning Across Tabular, NLP, and Vision Tasks

ApplicationsDGX agent

arXiv:2605.24340v1 Announce Type: new Abstract: Production deep learning systems across enterprise domains operate under constraints that academic benchmarks routinely obscure: labeled data is expensi

Contractual Skills: A GovernSpec Design Framework for Enterprise AI Agents

SafetyDGX agent

arXiv:2605.22634v2 Announce Type: replace-cross Abstract: Skills have become a practical packaging mechanism for agent instructions, workflows, scripts, and reference materials. In enterprise settings

Convex-Neural RRT*: Fast and Reliable Learning-Guided Sampling for High-Quality Robot Path Planning

Model ReleasesDGX agent

arXiv:2605.25006v1 Announce Type: cross Abstract: Sampling-based algorithms for robot path planning offer probabilistic completeness and strong empirical convergence properties across environments wit

CopulaSMOTE: A Copula-Based Oversampling Approach for Imbalanced Classification in Diabetes Prediction

Local AiDGX agent

arXiv:2506.17326v3 Announce Type: replace Abstract: Class imbalance remains a practical obstacle in the development of clinical prediction models for conditions such as diabetes mellitus, where the nu

CoRe-Code: Collaborative Reinforcement Learning for Code Generation

Local AiDGX agent

arXiv:2605.24812v1 Announce Type: new Abstract: Large language models (LLMs) have achieved strong performance in code generation, but most methods rely on autoregressive decoding without global planni

CP-Agent: A Calibrated Risk-Controlled Agent for Feedback-Driven Competitive Programming

AgentsDGX agent

arXiv:2605.24693v1 Announce Type: new Abstract: Large language models still struggle with contest-level programming, while many agentic remedies rely on massive inference-time sampling or expensive mu

DarkForest: Less Talk, Higher Accuracy for Multi-Agent LLMs

Model ReleasesDGX agent

arXiv:2605.25188v1 Announce Type: new Abstract: Multi-agent LLM systems improve reasoning by combining outputs from multiple agents, but interaction-heavy methods can introduce error propagation and h

Data-Specific Hyper-Parameter Design: A Paradigm Shift in Reservoir Computing

Model ReleasesDGX agent

arXiv:2605.25221v1 Announce Type: cross Abstract: Reservoir computing typically relies on large, randomly generated reservoirs, enabling simple, often linear readouts. Over the past two decades, most

DBPnet: Damper Characteristics-Based Bayesian Physics-Informed Neural Network for Wheel Load Estimation

SafetyDGX agent

arXiv:2605.24860v1 Announce Type: cross Abstract: Advanced driver assistance systems (ADAS) play an important role in modern automotive intelligence, significantly enhancing vehicle safety and stabili

Deep ZakaiJ: Structured Filtering for Jump-Diffusion Time Series Forecasting

ResearchDGX agent

arXiv:2605.24548v1 Announce Type: new Abstract: Time series driven by unobserved latent states frequently exhibit abrupt jump discontinuities whose timing and magnitude cannot be predicted from observ

DTO: a Differentiable Training Objective for Effective Counterfactual Story Rewriting

Local AiDGX agent

arXiv:2605.24885v1 Announce Type: new Abstract: Counterfactual story rewriting is a natural language processing task that requires updating an existing story to reflect a chosen alternative event, yet

DUEL: Adversarial Self-Play for Multimodal Reasoning

ResearchDGX agent

arXiv:2605.24794v1 Announce Type: cross Abstract: Reinforcement learning (RL) has emerged as an effective paradigm for improving the reasoning capability of vision-language models (VLMs). However, RL-

Dynamics Reveals Structure: Challenging the Linear Propagation Assumption

Model ReleasesDGX agent

arXiv:2601.21601v2 Announce Type: replace-cross Abstract: Neural networks adapt through first-order parameter updates, yet it remains unclear whether such updates preserve logical coherence. We invest

ECo-MoE: Embodiment-Conditioned Mixture of Experts Increases the Evolvability of Robots

SafetyDGX agent

arXiv:2605.24225v1 Announce Type: new Abstract: In this paper, we introduce a model of evolution and learning in robots that co-optimizes a distribution of latent design vectors (genotypes) and a mixt

EfficientGraph-RAG: Structured Retrieval-State Management for Cross-Task Retrieval-Augmented Generation

AgentsDGX agent

arXiv:2605.25379v1 Announce Type: new Abstract: Retrieval-augmented generation (RAG) has become the standard way to ground large language models in external knowledge, but many systems still organize

Estimating Mixture Distributions via Stochastic Mirror Descent

ResearchDGX agent

arXiv:2605.24929v1 Announce Type: cross Abstract: We revisit the classical problem of estimating an unknown distribution from its samples by fitting a mixture model that minimizes cross-entropy loss.

EvoCode-Bench: Evaluating Coding Agents in Multi-Turn Iterative Interactions

Model ReleasesDGX agent

arXiv:2605.24110v1 Announce Type: new Abstract: Coding agents are increasingly used as iterative development partners, but most benchmarks still evaluate one specification followed by one final assess

EvoSci: A Bio-Inspired Multi-Agent Framework for the Evolution of Scientific Discovery

AgentsDGX agent

arXiv:2605.24018v1 Announce Type: new Abstract: Large language models (LLMs), have shown strong potential in scientific discovery, yet existing methods still face substantial challenges in the design

Explainable Retinal Imaging for Prediction of Multi-Organ Dysfunction in Type 2 Diabetes

ResearchDGX agent

arXiv:2605.24912v1 Announce Type: cross Abstract: Background: Type 2 diabetes mellitus (T2DM) is increasingly recognised as a systemic disease characterised by coordinated dysfunction across metabolic

Exploration of Perceptual Speech Features for Clinical Decision-Support in Mental Health Care

Model ReleasesDGX agent

arXiv:2605.24678v1 Announce Type: new Abstract: Speech and language technologies offer valuable opportunities for supporting mental health assessment through objective and interpretable cues. We prese

Extreme Region Policy Distillation

SafetyDGX agent

arXiv:2605.25582v1 Announce Type: cross Abstract: Reinforcement learning for large language models faces a fundamental trade-off between sample efficiency and asymptotic performance: strictly on-polic

Faithful or Fabricated? A Causal Framework for Rationalization Bias in LLM Judges

SafetyDGX agent

arXiv:2605.23970v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used as automatic judges for summarization and dialogue evaluation. Prior work has documented biases such

Feature Learning in Wide Neural Networks under muP: Identifiability and Sparse-Dictionary Decomposition of the Mean-Field Limit

Model ReleasesDGX agent

arXiv:2605.24710v1 Announce Type: new Abstract: We establish four structural results for feature learning in wide two-layer neural networks under the Maximal Update Parametrization (muP). First, we pr

Federated Sketching LoRA: A Flexible Framework for Heterogeneous Collaborative Fine-Tuning of LLMs

ResearchDGX agent

arXiv:2501.19389v4 Announce Type: replace Abstract: Fine-tuning large language models (LLMs) on resource-constrained clients remains a challenging problem. Recent works have fused low-rank adaptation

From DPPs to k-DPPs: identifiability analysis via spectral decomposition

Model ReleasesDGX agent

arXiv:2605.25526v1 Announce Type: cross Abstract: We study the geometry of determinantal point processes (DPPs) through the spectral decomposition L=ULambda U^{op}. The spectrum Lambda governs the car

From Reasoning to Code: GRPO Optimization for Underrepresented Languages

SafetyDGX agent

arXiv:2506.11027v3 Announce Type: replace-cross Abstract: Generating accurate and executable code using Large Language Models (LLMs) remains a significant challenge for underrepresented programming la

GDformer: Going Beyond Subsequence Isolation for Multivariate Time Series Anomaly Detection

Model ReleasesDGX agent

arXiv:2501.18196v3 Announce Type: replace Abstract: Unsupervised anomaly detection of multivariate time series is a challenging task, given the requirements of deriving a compact detection criterion w

Generative Neural Operators through Diffusion Last Layer

ResearchDGX agent

arXiv:2602.04139v2 Announce Type: replace Abstract: Neural operators provide a powerful framework for learning discretization invariant mappings between function spaces, but standard deterministic mod

GreenSeg: Ground Segmentation Algorithm for Agricultural Robots in Mediterranean Greenhouses using RGB-D Point Clouds

Model ReleasesDGX agent

arXiv:2605.25279v1 Announce Type: new Abstract: Greenhouse agriculture in the Mediterranean region faces significant automation challenges due to its unique structural and environmental constraints. T

Harnessing AtomisticSkills for Agentic Atomistic Research

AgentsDGX agent

arXiv:2605.24002v1 Announce Type: cross Abstract: Computational materials science and chemistry span vast knowledge domains and fractured software ecosystems. Although large language models (LLMs) hav

Hide to Guide: Learning via Semantic Masking

SafetyDGX agent

arXiv:2605.25198v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has become a powerful paradigm for improving language models on reasoning-intensive tasks, but i

How Should LLMs Consume High-Quality Data? Optimal Data Scheduling via Quality-Aware Functional Scaling Laws

TutorialsDGX agent

arXiv:2605.25698v1 Announce Type: cross Abstract: High-quality data is scarce in large language model (LLM) training, yet how to schedule its use jointly with training dynamics lacks theoretical guida

Huge one for developers building with AI. Excited to announce @ClementDelangue (CEO @huggingface) is joining DASH 2026 for a fireside chat w…

ApplicationsDGX agent

Huge one for developers building with AI. Excited to announce @ClementDelangue (CEO @huggingface) is joining DASH 2026 for a fireside chat with @oliveur. Open source changed software, open models are

Hybrid Deep Searcher: Scalable Parallel and Sequential Search Reasoning

AgentsDGX agent

arXiv:2508.19113v3 Announce Type: replace Abstract: Large reasoning models (LRMs) combined with retrieval-augmented generation (RAG) have enabled deep research agents capable of multi-step reasoning w

I don't comment on every article that uses out-of-date measures on AI ability, but I felt (probably wrongly) that the article was a response…

Model ReleasesDGX agent

I don't comment on every article that uses out-of-date measures on AI ability, but I felt (probably wrongly) that the article was a response to my viral tweet, so I felt I needed to say something! htt

Identifying and Mitigating Systemic Measurement Bias in Production LLM Inference Benchmarks

SafetyDGX agent

arXiv:2605.24217v1 Announce Type: new Abstract: As Large Language Models (LLMs) transition from research environments to production deployments, evaluating their performance against strict Service Lev

Intent Signal Theory: A Computational Framework for Intent-State Control in Human-AI Interaction

ResearchDGX agent

arXiv:2605.25058v1 Announce Type: cross Abstract: Current AI interaction models treat the prompt as the primary object of exchange, omitting a critical layer: the user's latent source intent, the goal

Interdomain Attention: Beyond Token-Level Key-Value Memory

ResearchDGX agent

arXiv:2605.24330v1 Announce Type: new Abstract: Transformers and deep state space models (SSMs) sit at opposite ends of a basic design choice: attention routes each query through a growing key-value (

Interpretable and backpropagation-free Green Learning for efficient multi-task echocardiographic segmentation and classification

ResearchDGX agent

arXiv:2601.19743v3 Announce Type: replace-cross Abstract: Echocardiography is a cornerstone for managing heart failure (HF), with Left Ventricular Ejection Fraction (LVEF) being a critical metric for

Interpretation, Learning, and Empathy as One Constraint: A Residual-Adequacy Architecture with Accountable Abstention

Local AiDGX agent

arXiv:2605.24999v1 Announce Type: cross Abstract: An agent must act on the situation before it, learn what it cannot yet represent, and model other agents well enough to coordinate. These faculties ar

Introducing CHI-Bench on @huggingface: the world’s first long-horizon healthcare benchmark for AI agents. 75 real healthcare workflows + 20 …

Model ReleasesDGX agent

Introducing CHI-Bench on @huggingface: the world’s first long-horizon healthcare benchmark for AI agents. 75 real healthcare workflows + 20 apps + 200+ MCP tools + 1,290 skills + process / outcome rew

Is SaaS dead?

AgentsDGX agent

This article examines whether the Software-as-a-Service (SaaS) business model remains viable and competitive in the current market landscape. It likely discusses challenges facing SaaS companies, such

IVR-R1: Refining Trajectories through Iterative Visual-Grounded Reasoning in Reinforcement Learning

SafetyDGX agent

arXiv:2605.23997v1 Announce Type: cross Abstract: Multimodal large language models via reinforcement learning (RL) have demonstrated remarkable capabilities in complex visual reasoning tasks, yet they

Joint Optimization of Training and Inference in Federated Edge Learning via Constrained Multi-Objective Deep Reinforcement Learning

SafetyDGX agent

arXiv:2605.25916v1 Announce Type: new Abstract: Federated edge learning (FEEL) has recently emerged as a promising paradigm for achieving edge intelligence (EI) via enabling collaborative model traini

JudgmentBench: Comparing Rubric and Preference Evaluation for Quality Assessment

Model ReleasesDGX agent

arXiv:2605.25240v1 Announce Type: cross Abstract: Two methodologies dominate current practices of benchmarking: rubric-based scoring evaluates items against predefined criteria, whereas comparative ju

Knowing but Not Showing: LLMs Recognize Ambiguity but Rarely Ask Clarifying Questions

ResearchDGX agent

arXiv:2605.25284v1 Announce Type: new Abstract: User queries are often underspecified and may admit multiple valid interpretations. Rather than silently making assumptions about the user's intent, a h

Knowledge Graph-Driven Expert-Level Reasoning for Neuroscience

ResearchDGX agent

arXiv:2605.25183v1 Announce Type: cross Abstract: Knowledge graph (KG) is an abstraction that can be extracted from text corpora and used for in-depth reasoning. Prior work has leveraged KGs to fine-t

KT4EQG: Personalized Exercise Question Generation via Knowledge Tracing

ResearchDGX agent

arXiv:2605.23933v1 Announce Type: cross Abstract: Educational Question Generation (EQG) aims to synthesize customized exercise questions that enhance student learning. An effective EQG system should i

Latent Q-Barrier Shielding for Safe In-Context Reinforcement Learning

Model ReleasesDGX agent

arXiv:2605.25267v1 Announce Type: cross Abstract: Safe in-context reinforcement learning (ICRL) adapts online from interaction history without test-time parameter updates while controlling episode cos

LC-ERD: Mining Latent Logic for Self-Evolving Reasoning via Consistency-Regulated Reward Decomposition

SafetyDGX agent

arXiv:2605.24005v1 Announce Type: new Abstract: The evolution of Large Language Model (LLM) reasoning is bottlenecked by the scarcity of high-quality process data. While self-alignment via endogenous

Length Generalization with Log-Depth Recurrent Units

ResearchDGX agent

arXiv:2605.26035v1 Announce Type: new Abstract: Length generalization remains a persistent challenge for neural networks: recurrent models tend to suffer from positional biases, while transformers are

LETS Forecast: Learning Embedology for Time Series Forecasting

ApplicationsDGX agent

arXiv:2506.06454v2 Announce Type: cross Abstract: Real-world time series are often governed by complex nonlinear dynamics. Understanding these underlying dynamics is crucial for precise future predict

LGMT: Logic-Grounded Metamorphic Testing for Evaluating the Reasoning Reliability of LLMs

ResearchDGX agent

arXiv:2605.23965v1 Announce Type: new Abstract: Large Language Models (LLMs) achieve strong performance on logical reasoning benchmarks, yet their reliability remains uncertain. Existing evaluations r

LLM Agent Based Renewable Energy Forecasting Using Edge and IoT Data A Review of Solar Wind Weather and Grid Aware Decision Support

Local AiDGX agent

arXiv:2605.25141v1 Announce Type: cross Abstract: Reliable forecasting of renewable energy generation is a foundational requirement for grid stability energy trading battery scheduling and carbon awar

Localization then Neutralization: Gradient-guided Token Suppression against Visual Prompt Injection Attack

Local AiDGX agent

arXiv:2605.25194v1 Announce Type: new Abstract: Adversarial images pose a severe security threat to multimodal large language models through prompt injection. Existing defenses largely lack a principl

← Previous
1…665666667668669…1042
Next →