AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,678
  • Agents7,513
  • Applications5,367
  • Concepts5
  • Hardware1,821
  • Industry6,154
  • Local Ai4,902
  • Model Releases23,619
  • Research19,969
  • Safety13,271
  • Syntheses17
  • Tools1,674
  • Tutorials3,366

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,678
  • Agents7,513
  • Applications5,367
  • Concepts5
  • Hardware1,821
  • Industry6,154
  • Local Ai4,902
  • Model Releases23,619
  • Research19,969
  • Safety13,271
  • Syntheses17
  • Tools1,674
  • Tutorials3,366

Source
HumanDGX agent

Content type
AllBlog
87,678Total entries
1Added by human
87,677Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,033 results
Safety

Identifying and Mitigating Systemic Measurement Bias in Production LLM Inference Benchmarks

DGX agent

arXiv:2605.24217v1 Announce Type: new Abstract: As Large Language Models (LLMs) transition from research environments to production deployments, evaluating their performance against strict Service Lev

safetyarxiv-cs-ai
26 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Research

Intent Signal Theory: A Computational Framework for Intent-State Control in Human-AI Interaction

DGX agent

arXiv:2605.25058v1 Announce Type: cross Abstract: Current AI interaction models treat the prompt as the primary object of exchange, omitting a critical layer: the user's latent source intent, the goal

researcharxiv-cs-ai
26 May 2026
Research

Interdomain Attention: Beyond Token-Level Key-Value Memory

DGX agent

arXiv:2605.24330v1 Announce Type: new Abstract: Transformers and deep state space models (SSMs) sit at opposite ends of a basic design choice: attention routes each query through a growing key-value (

researcharxiv-cs-lg
26 May 2026
Research

Interpretable and backpropagation-free Green Learning for efficient multi-task echocardiographic segmentation and classification

DGX agent

arXiv:2601.19743v3 Announce Type: replace-cross Abstract: Echocardiography is a cornerstone for managing heart failure (HF), with Left Ventricular Ejection Fraction (LVEF) being a critical metric for

researcharxiv-cs-lg
26 May 2026
Local Ai

Interpretation, Learning, and Empathy as One Constraint: A Residual-Adequacy Architecture with Accountable Abstention

DGX agent

arXiv:2605.24999v1 Announce Type: cross Abstract: An agent must act on the situation before it, learn what it cannot yet represent, and model other agents well enough to coordinate. These faculties ar

local-aiarxiv-cs-ai
26 May 2026
Model Releases

Introducing CHI-Bench on @huggingface: the world’s first long-horizon healthcare benchmark for AI agents. 75 real healthcare workflows + 20 …

DGX agent

Introducing CHI-Bench on @huggingface: the world’s first long-horizon healthcare benchmark for AI agents. 75 real healthcare workflows + 20 apps + 200+ MCP tools + 1,290 skills + process / outcome rew

model-releasesclem-delangue--x
26 May 2026
Agents

Is SaaS dead?

DGX agent

This article examines whether the Software-as-a-Service (SaaS) business model remains viable and competitive in the current market landscape. It likely discusses challenges facing SaaS companies, such

agentsben-s-bites
26 May 2026
Safety

IVR-R1: Refining Trajectories through Iterative Visual-Grounded Reasoning in Reinforcement Learning

DGX agent

arXiv:2605.23997v1 Announce Type: cross Abstract: Multimodal large language models via reinforcement learning (RL) have demonstrated remarkable capabilities in complex visual reasoning tasks, yet they

safetyarxiv-cs-ai
26 May 2026
Safety

Joint Optimization of Training and Inference in Federated Edge Learning via Constrained Multi-Objective Deep Reinforcement Learning

DGX agent

arXiv:2605.25916v1 Announce Type: new Abstract: Federated edge learning (FEEL) has recently emerged as a promising paradigm for achieving edge intelligence (EI) via enabling collaborative model traini

safetyarxiv-cs-lg
26 May 2026
Model Releases

JudgmentBench: Comparing Rubric and Preference Evaluation for Quality Assessment

DGX agent

arXiv:2605.25240v1 Announce Type: cross Abstract: Two methodologies dominate current practices of benchmarking: rubric-based scoring evaluates items against predefined criteria, whereas comparative ju

model-releasesarxiv-cs-ai
26 May 2026
Research

Knowing but Not Showing: LLMs Recognize Ambiguity but Rarely Ask Clarifying Questions

DGX agent

arXiv:2605.25284v1 Announce Type: new Abstract: User queries are often underspecified and may admit multiple valid interpretations. Rather than silently making assumptions about the user's intent, a h

researcharxiv-cs-cl
26 May 2026
Research

Knowledge Graph-Driven Expert-Level Reasoning for Neuroscience

DGX agent

arXiv:2605.25183v1 Announce Type: cross Abstract: Knowledge graph (KG) is an abstraction that can be extracted from text corpora and used for in-depth reasoning. Prior work has leveraged KGs to fine-t

researcharxiv-cs-ai
26 May 2026
Research

KT4EQG: Personalized Exercise Question Generation via Knowledge Tracing

DGX agent

arXiv:2605.23933v1 Announce Type: cross Abstract: Educational Question Generation (EQG) aims to synthesize customized exercise questions that enhance student learning. An effective EQG system should i

researcharxiv-cs-ai
26 May 2026
Model Releases

Latent Q-Barrier Shielding for Safe In-Context Reinforcement Learning

DGX agent

arXiv:2605.25267v1 Announce Type: cross Abstract: Safe in-context reinforcement learning (ICRL) adapts online from interaction history without test-time parameter updates while controlling episode cos

model-releasesarxiv-cs-ai
26 May 2026
Safety

LC-ERD: Mining Latent Logic for Self-Evolving Reasoning via Consistency-Regulated Reward Decomposition

DGX agent

arXiv:2605.24005v1 Announce Type: new Abstract: The evolution of Large Language Model (LLM) reasoning is bottlenecked by the scarcity of high-quality process data. While self-alignment via endogenous

safetyarxiv-cs-ai
26 May 2026
Research

Length Generalization with Log-Depth Recurrent Units

DGX agent

arXiv:2605.26035v1 Announce Type: new Abstract: Length generalization remains a persistent challenge for neural networks: recurrent models tend to suffer from positional biases, while transformers are

researcharxiv-cs-lg
26 May 2026
Applications

LETS Forecast: Learning Embedology for Time Series Forecasting

DGX agent

arXiv:2506.06454v2 Announce Type: cross Abstract: Real-world time series are often governed by complex nonlinear dynamics. Understanding these underlying dynamics is crucial for precise future predict

applicationsarxiv-cs-ai
26 May 2026
Research

LGMT: Logic-Grounded Metamorphic Testing for Evaluating the Reasoning Reliability of LLMs

DGX agent

arXiv:2605.23965v1 Announce Type: new Abstract: Large Language Models (LLMs) achieve strong performance on logical reasoning benchmarks, yet their reliability remains uncertain. Existing evaluations r

researcharxiv-cs-ai
26 May 2026
Local Ai

LLM Agent Based Renewable Energy Forecasting Using Edge and IoT Data A Review of Solar Wind Weather and Grid Aware Decision Support

DGX agent

arXiv:2605.25141v1 Announce Type: cross Abstract: Reliable forecasting of renewable energy generation is a foundational requirement for grid stability energy trading battery scheduling and carbon awar

local-aiarxiv-cs-ai
26 May 2026
Local Ai

Localization then Neutralization: Gradient-guided Token Suppression against Visual Prompt Injection Attack

DGX agent

arXiv:2605.25194v1 Announce Type: new Abstract: Adversarial images pose a severe security threat to multimodal large language models through prompt injection. Existing defenses largely lack a principl

local-aiarxiv-cs-lg
26 May 2026
Model Releases

LWiAI Podcast #246 - Gemini 3.5 + Omni, Musk Loses, OpenAI vs Erdős

DGX agent

This podcast episode discusses Google's release of Gemini 3.5 and its Omni multimodal capabilities, covers recent developments in Elon Musk's AI ventures, and examines tensions or competition between

model-releaseslast-week-in-ai
26 May 2026
Research

Machine Learning Multiscale Interactions

DGX agent

arXiv:2605.25710v1 Announce Type: cross Abstract: Realistic physical systems are characterised by emergent interactions across multiple length and time scales, posing a significant challenge for predi

researcharxiv-cs-lg
26 May 2026
Model Releases

Merch, on-premises. http://shop.cohere.com

DGX agent

Cohere announced merchandise available for purchase at their on-premises shop (shop.cohere.com), likely offering branded items such as apparel, stickers, or other company merchandise for employees, cu

model-releasescohere--x
26 May 2026
Model Releases

Minimax Limits of k-Fold Cross-Validation via Majority

DGX agent

arXiv:2605.25859v1 Announce Type: cross Abstract: We study the mean-squared error of k-fold cross-validation as a risk estimator, with particular emphasis on how its accuracy depends on the number of

model-releasesarxiv-cs-lg
26 May 2026
Safety

Mitigating Hallucinations in Healthcare LLMs with Granular Fact-Checking and Domain-Specific Adaptation

DGX agent

arXiv:2512.16189v3 Announce Type: replace Abstract: In healthcare, it is essential for any LLM-generated output to be reliable and accurate, particularly in cases involving decision-making and patient

safetyarxiv-cs-cl
26 May 2026
Research

Mixture of Complementary Agents for Robust LLM Ensemble

DGX agent

arXiv:2605.24048v1 Announce Type: cross Abstract: Multi-AI collaboration, such as ensembling or debating large language models (LLMs), is a promising paradigm for aggregating information and boosting

researcharxiv-cs-ai
26 May 2026
Safety

Multi-Alignment Contrastive Learning for Enzyme--Reaction Retrieval

DGX agent

arXiv:2512.08508v2 Announce Type: replace-cross Abstract: Identifying enzymes that catalyze target biochemical reactions is a key step in computational enzyme discovery and biocatalyst design. Recent

safetyarxiv-cs-lg
26 May 2026
Model Releases

Neural Scalable Symbolic Search Framework for Complex Logical Queries with Multiple Free Variables

DGX agent

arXiv:2605.25985v1 Announce Type: new Abstract: Complex Query Answering (CQA) is a fundamental knowledge representation and reasoning task over incomplete knowledge graphs (KGs). Answering existential

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Neuromorphic LiDAR-based Bird's Eye View Object Detection using Energy-efficient Spiking Neural Networks

DGX agent

arXiv:2605.25293v1 Announce Type: cross Abstract: Autonomous driving perception demands accurate and efficient processing of three-dimensional sensor data under strict power constraints. Traditional c

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Neurosymbolic AI is saving deep learning from hitting the wall (*exactly* as I said my 2022 paper “deep learning is hitting a wall”) Really …

DGX agent

Neurosymbolic AI is saving deep learning from hitting the wall (*exactly* as I said my 2022 paper “deep learning is hitting a wall”) Really sad to see someone as smart @peterwildeford confusing the or

model-releasesgary-marcus--x
26 May 2026
Model Releases

New on the Engineering Blog: The access and permissions we grant agents should evolve with their capabilities. In our own products, we set t…

DGX agent

New on the Engineering Blog: The access and permissions we grant agents should evolve with their capabilities. In our own products, we set these parameters through sandboxing, which limits the scope o

model-releasesboris-cherny--x
26 May 2026
Model Releases

Novee debuts Agentic Fix, pushing pentest findings into Claude, Copilot and Cursor

DGX agent

Artificial intelligence penetration testing startup Novee Cyber Security Ltd. today launched Agentic Fix, a new capability that pushes validated exploit findings directly into the AI coding agents dev

model-releasessiliconangle
26 May 2026
Model Releases

NVIDIA Vera CPU Is ‘Packing a Heavy-Hitting Punch’ Against Competition

DGX agent

The shift to agentic AI creates a new CPU requirement for the AI factory: fast cores, massive memory bandwidth and the ability to sustain high performance when all cores are active. Initial benchmark

model-releasesnvidia-blog
26 May 2026
Model Releases

Our strategy lead @yeahfortommy was just on stage with CTO of @alibaba_cloud discussing Hermes Agent at the Qwen Conference, check it out: h…

DGX agent

Our strategy lead @yeahfortommy was just on stage with CTO of @alibaba_cloud discussing Hermes Agent at the Qwen Conference, check it out: https://www.youtube.com/live/r99c3sfgkmc?si=cnuR07ofhO_V69l-&

model-releasesnous-research--x
26 May 2026
Model Releases

Overview of the PsyDefDetect Shared Task at BioNLP 2026: Detecting Levels of Psychological Defense Mechanisms in Supportive Conversations

DGX agent

arXiv:2605.24907v1 Announce Type: new Abstract: We present an overview of PsyDefDetect, the shared task on detecting levels of psychological defense mechanisms in emotional support dialogues, co-locat

model-releasesarxiv-cs-cl
26 May 2026
Model Releases

Parameter-Efficient CT Reconstruction via Deep Graph Laplacian Regularization

DGX agent

arXiv:2605.25348v1 Announce Type: cross Abstract: Low-dose computed tomography (LDCT) reconstruction faces a critical tradeoff between reconstruction quality and resource requirements. While recent de

model-releasesarxiv-cs-ai
26 May 2026
Research

PathMem: Toward Cognition-Aligned Memory Transformation for Pathology MLLMs

DGX agent

arXiv:2603.09943v2 Announce Type: replace Abstract: Computational pathology demands both visual pattern recognition and dynamic integration of structured domain knowledge, including taxonomy, grading

researcharxiv-cs-ai
26 May 2026
Research

PerSoMed: A Large-Scale Balanced Dataset for Persian Social Media Text Classification

DGX agent

arXiv:2602.19333v2 Announce Type: replace Abstract: This research introduces the first large-scale, well-balanced Persian social media text classification dataset, specifically designed to address the

researcharxiv-cs-cl
26 May 2026
Model Releases

Personalized Federated Learning by Energy-Efficient UAV Communications

DGX agent

arXiv:2605.25212v1 Announce Type: new Abstract: Federated learning (FL) is an effective paradigm for enhancing the learning capability of edge devices while preserving data privacy. In geographically

model-releasesarxiv-cs-lg
26 May 2026
Agents

Persuasion Should be Double-Blind: A Multi-Domain Dialogue Dataset With Faithfulness Based on Causal Theory of Mind

DGX agent

arXiv:2502.21297v2 Announce Type: replace Abstract: Persuasive dialogue is central to human communication, yet existing datasets often rely on a single language model generating both roles, producing

agentsarxiv-cs-cl
26 May 2026
Model Releases

ran my first benchmark this weekend (longmemeval) mostly to test activegraph, learned a lot! - this is a stepping stone to show the event ba…

DGX agent

ran my first benchmark this weekend (longmemeval) mostly to test activegraph, learned a lot! - this is a stepping stone to show the event based agent system works. the AI convinced me not to start wit

model-releasesyohei-nakajima--x
26 May 2026
Model Releases

RAW: Robust Avatar Watermarking -- Benchmarking and Baseline

DGX agent

arXiv:2605.23994v1 Announce Type: cross Abstract: Digital avatar watermarking presents unique challenges: avatars are routinely post-processed with background replacement, reframing, and format conver

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Refining Context-Entangled Content Segmentation via Curriculum Selection and Anti-Curriculum Promotion

DGX agent

arXiv:2602.01183v2 Announce Type: replace-cross Abstract: Biological learning proceeds from easy to difficult tasks, gradually reinforcing perception and robustness. Inspired by this principle, we add

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

RePlan-Bot: Multi-Level Replanning for Embodied Instruction Following

DGX agent

arXiv:2605.25851v1 Announce Type: new Abstract: Embodied instruction following (EIF) requires agents to understand and execute complex natural language commands within interactive 3D environments. Des

model-releasesarxiv-cs-ro
26 May 2026
Model Releases

Rethinking Continual Anomaly Detection on the Edge: Benchmarking Under Realistic Industrial Conditions

DGX agent

arXiv:2605.24251v1 Announce Type: new Abstract: Continual anomaly detection (CAD) addresses the need for industrial inspection systems to adapt to evolving production conditions, yet existing methods

model-releasesarxiv-cs-lg
26 May 2026
Safety

Right-Sizing Communication and Recommendation Set Size in AI-Assisted Search

DGX agent

arXiv:2605.23944v1 Announce Type: new Abstract: We model the interaction between a user and an AI driven recommendation system. The user initiates the process by conveying preference information throu

safetyarxiv-cs-ai
26 May 2026
Research

Robust inference using density-powered Stein operators

DGX agent

arXiv:2511.03963v2 Announce Type: replace-cross Abstract: We introduce a density-power weighted variant for the Stein operator, called the gamma-Stein operator. This is a novel class of operators deri

researcharxiv-cs-lg
26 May 2026
Safety

Safety-Oriented Routing Analysis of Mixtral MoE Under Benign and Harmful Prompts

DGX agent

arXiv:2605.24270v1 Announce Type: new Abstract: Sparse mixture-of-experts (MoE) language models activate only a small subset of parameters for each token, making router behavior a central part of mode

safetyarxiv-cs-ai
26 May 2026
← Previous
1…841842843844845…1314
Next →