AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries90,914
  • Agents7,747
  • Applications5,533
  • Concepts5
  • Hardware1,920
  • Industry6,196
  • Local Ai5,094
  • Model Releases24,727
  • Research20,781
  • Safety13,739
  • Syntheses17
  • Tools1,678
  • Tutorials3,477

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries90,914
  • Agents7,747
  • Applications5,533
  • Concepts5
  • Hardware1,920
  • Industry6,196
  • Local Ai5,094
  • Model Releases24,727
  • Research20,781
  • Safety13,739
  • Syntheses17
  • Tools1,678
  • Tutorials3,477

Source
HumanDGX agent

Content type
90,914Total entries
1Added by human
90,913Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
90,913 results
Safety

Steering Beyond the Support: Adversarial Training on Unsupervised Jailbroken Activation Simulation

DGX agent

arXiv:2605.24535v1 Announce Type: cross Abstract: Jailbreak prompts can trigger harmful completions on aligned LLMs, In accordance, safety steering has been proposed: test-time activation intervention

safetyarxiv-cs-lg
26 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Stein Variational Ergodic Surface Coverage with SE(3) Constraints

DGX agent

arXiv:2603.09458v3 Announce Type: replace Abstract: Surface manipulation tasks require robots to generate trajectories that comprehensively cover complex 3D surfaces while maintaining precise end-effe

model-releasesarxiv-cs-ro
26 May 2026
Research

Step-TP: A Grounded, Step-Level Dataset with Chain-of-Thought Reasoning for LLM-Guided Tensor Program Optimization

DGX agent

arXiv:2605.25954v1 Announce Type: cross Abstract: Despite the strong reasoning capabilities of large language models (LLMs), optimizing the execution efficiency of tensor programs remains challenging

researcharxiv-cs-ai
26 May 2026
Research

StepGap: A Hybrid NLI-LLM Checker for Step-Level Evidence-Gap Detectionin Multi-Hop Question Answering

DGX agent

arXiv:2605.24733v1 Announce Type: new Abstract: We present extbf{StepGap}, a hybrid NLI-LLM decision tree that detects step-level evidence gaps in multi-hop QA and emits one of three typed labels: ext

researcharxiv-cs-cl
26 May 2026
Research

Stiffness Optimization for Concentrated Bending in Magnetically Actuated Catheters: Maintaining Steerability under Gradient Stiffness

DGX agent

arXiv:2605.25005v1 Announce Type: new Abstract: Achieving both efficient pushability (propulsion transmission) and proximally concentrated bending for steerability is challenging for magnetically actu

researcharxiv-cs-ro
26 May 2026
Model Releases

Stochastic Estimation of the Layer-wise Hessian Trace for Monitoring Neural-network Training

DGX agent

arXiv:2605.25674v1 Announce Type: new Abstract: The loss and the norm of its gradient separate the healthy and the pathological regimes of neural-network training only weakly, whilst the curvature of

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

Stochastic Linear Bandits with Parameter Noise

DGX agent

arXiv:2601.23164v2 Announce Type: replace Abstract: We study the stochastic linear bandits with parameter noise model, in which the reward of action a is a^op heta where heta is sampled i.i.d. We show

model-releasesarxiv-cs-lg
26 May 2026
Safety

Stop Comparing LLM Agents Without Disclosing the Harness

DGX agent

arXiv:2605.23950v1 Announce Type: new Abstract: This position paper argues that, for long-horizon tasks evaluated across models with comparable frontier capability, the agent execution harness, namely

safetyarxiv-cs-ai
26 May 2026
Agents

STORM: Internalized Modeling for Spatial-Temporal Reasoning in Video-Language Models

DGX agent

arXiv:2605.26014v1 Announce Type: cross Abstract: Many video reasoning tasks require tracking motion, temporal order, and evolving visual states across frames. Existing methods built on large vision-l

agentsarxiv-cs-cl
26 May 2026
Safety

Strat-Reasoner: Reinforcing Strategic Reasoning of LLMs in Multi-Agent Games

DGX agent

arXiv:2605.04906v2 Announce Type: replace Abstract: While Large Language Models (LLMs) excel in certain reasoning tasks, they struggle in multi-agent games where the final outcome depends on the joint

safetyarxiv-cs-ai
26 May 2026
Model Releases

STREAM: A Data-Centric Framework for Mining High-Value Task-Oriented Dialogues from Streaming Media

DGX agent

arXiv:2605.25162v1 Announce Type: cross Abstract: Large language models for vertical domains are bottlenecked by the scarcity of complex, domain-specific task-oriented dialogues. Existing data acquisi

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Streaming Reinforcement Learning under Partial Observability with Real-Time Recurrent Learning

DGX agent

arXiv:2605.24709v1 Announce Type: new Abstract: Streaming reinforcement learning has emerged as an online learning paradigm that conforms to the restrictions of natural learning agents that process da

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

StreamProfileBench: A Benchmark for Fine-Grained User Profile Inference in Real-World Streaming Scenarios

DGX agent

arXiv:2605.25758v1 Announce Type: new Abstract: Large Language Models (LLMs) have reshaped user profiling, yet current evaluations mainly focus on static data snapshots. This paradigm overlooks the re

model-releasesarxiv-cs-cl
26 May 2026
Applications

StrTransformer: Source-Wise Structured Transformers for Unsupervised Blind Source Recovery

DGX agent

arXiv:2605.25648v1 Announce Type: cross Abstract: This paper proposes StrTransformer, a source-wise structured Transformer framework for blind source recovery and branch-wise latent modeling. Instead

applicationsarxiv-cs-lg
26 May 2026
Model Releases

StructBreak: Structural Cognitive Overload-Induced Safety Failures in MLLMs

DGX agent

arXiv:2605.25534v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) excel at structural reasoning yet suffer from a sharp logical brittleness in structural consistency. We term th

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Structural Abstraction as an Inductive Bias for Non-Stationary Language Model Training

DGX agent

arXiv:2603.17198v2 Announce Type: replace-cross Abstract: A foundational principle in cognitive science holds that intelligent agents do not learn by storing experiences as isolated instances, but by

model-releasesarxiv-cs-cl
26 May 2026
Applications

Structure-Aware RAG: Structured Retrieval Augmented Generation from Noisy Data for Conversational Agents

DGX agent

arXiv:2605.24366v1 Announce Type: new Abstract: Large Language Models (LLMs) have been widely adopted in conversational applications. However, their reliance on parametric knowledge limits reliability

applicationsarxiv-cs-cl
26 May 2026
Research

Subspace Aggregation Query and Index Generation for Multidimensional Resource Space Model

DGX agent

arXiv:2505.02129v3 Announce Type: replace-cross Abstract: Organizing large-scale resources in a multidimensional semantic space is an approach to efficiently managing and querying resources from diffe

researcharxiv-cs-ai
26 May 2026
Safety

Subspace-Guided Semantic and Topological Invariant Registration for Annotation-Free Ultrasound Plane Quality Control

DGX agent

arXiv:2605.25396v1 Announce Type: cross Abstract: Reliable quality control (QC) of ultrasound images is essential for both real-time acquisition guidance and retrospective clinical audit, yet existing

safetyarxiv-cs-ai
26 May 2026
Research

Sum of Costs Diffusion with Dynamic Guidance for Motion Planning

DGX agent

arXiv:2605.24690v1 Announce Type: cross Abstract: The motion planning problem for robotic manipulation can be addressed through classical or deep learning approaches. Existing methods face significant

researcharxiv-cs-lg
26 May 2026
Safety

Summoning the Oracle to Slay It: Mitigating Look-Ahead Bias in Financial Backtesting with Large Language Models

DGX agent

arXiv:2605.24564v1 Announce Type: new Abstract: Backtesting large language models (LLMs) on historical financial data is unreliable because pre-training cuts off after the events happened. An LLM trai

safetyarxiv-cs-ai
26 May 2026
Industry

Sundar Pichai on AI, the future of search, and what’s happening to the web

DGX agent

Today, I’m talking with Google and Alphabet CEO Sundar Pichai, in a conversation we recorded just after the Google I/O developer conference. This is the fifth year Sundar and I have sat down after I/O

industrythe-verge-ai
26 May 2026
Model Releases

SURGE: On the Potential of Large Language Models as General-Purpose Surrogate Code Executors

DGX agent

arXiv:2502.11167v5 Announce Type: replace-cross Abstract: Neural surrogate models are powerful and efficient tools in data mining. Meanwhile, large language models (LLMs) have demonstrated remarkable

model-releasesarxiv-cs-cl
26 May 2026
Research

Synheart Capacity: A Theory-Driven Physiological Representation of Cognitive Capacity Dynamics from Wearable Signals

DGX agent

arXiv:2605.24416v1 Announce Type: new Abstract: Human cognitive performance is constrained by limited mental resources, yet continuous computational estimation of cognitive capacity dynamics remains a

researcharxiv-cs-lg
26 May 2026
Model Releases

System scaling is the next real bottleneck in agentic AI. If you build agent orchestration layers, this is a clean map of where the engineer…

DGX agent

System scaling is the next real bottleneck in agentic AI. If you build agent orchestration layers, this is a clean map of where the engineering leverage actually sits. The labs own the model. You own

model-releasesdair-ai--x
26 May 2026
Research

T2S-MPC: Time-Embedded Online Adaptive Model Predictive Control for Time-Varying Dynamics

DGX agent

arXiv:2605.24852v1 Announce Type: new Abstract: Recent advances in learning-based model predictive control (MPC) have leveraged neural networks for online model learning, achieving strong performance

researcharxiv-cs-lg
26 May 2026
Research

TaBIIC2: Interactive Building of Ontological Taxonomies using Weighted Self-Organizing Maps

DGX agent

arXiv:2605.24899v1 Announce Type: new Abstract: Ontologies represent the conceptual knowledge of a domain. At the core of an ontology is the taxonomy of concepts and subconcepts that represent specifi

researcharxiv-cs-ai
26 May 2026
Safety

TapSampling: Inference-Time Sampling with a Task-Progress-Understanding Verifier for Robotic Manipulation

DGX agent

arXiv:2605.25547v1 Announce Type: new Abstract: Existing embodied control research demonstrates remarkable performance improvements by scaling training data and model size. We instead explore inferenc

safetyarxiv-cs-ro
26 May 2026
Safety

Task-Aligned Self-Supervised Learning for Medical Image Analysis: A Systematic Review and Practical Design Guidelines

DGX agent

arXiv:2605.23995v1 Announce Type: cross Abstract: Self-supervised learning (SSL) has emerged as a promising paradigm for addressing the annotation bottleneck in medical imaging by learning representat

safetyarxiv-cs-ai
26 May 2026
Model Releases

Teaching large language models to reason like expert diagnosticians

DGX agent

arXiv:2509.12194v2 Announce Type: replace Abstract: Differential diagnosis is an iterative process that integrates patient information with broader medical knowledge. Clinical case series such as the

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Teaching Through Analogies: A Modular Pipeline for Educational Analogy Generation

DGX agent

arXiv:2605.24211v1 Announce Type: cross Abstract: Analogies help learners understand unfamiliar concepts by relating them to known concepts. Despite recent advances, large language models (LLMs) conti

model-releasesarxiv-cs-ai
26 May 2026
Agents

Technical deep dive: AgentCore payments and innovation in agentic commerce

DGX agent

Amazon Bedrock AgentCore payments is now available in preview, it provides instant payments to paid external services with no manual billing setup per provider, stablecoin support for cost-effective m

agentsaws-ml-blog
26 May 2026
Applications

Temporal Concept Drift in Legal Judgment Prediction: Neural Baselines Across Three Epochs of Ukrainian Court Decisions

DGX agent

arXiv:2605.24452v1 Announce Type: cross Abstract: Legal NLP benchmarks evaluate models on randomly split data, implicitly assuming that legal language is stationary. We test this assumption by fine-tu

applicationsarxiv-cs-ai
26 May 2026
Local Ai

Temporal Score Rescaling for Temperature Sampling in Diffusion and Flow Models

DGX agent

arXiv:2510.01184v2 Announce Type: replace Abstract: We present a mechanism to steer the sampling diversity of denoising diffusion and flow matching models, allowing users to sample from a sharper or b

local-aiarxiv-cs-lg
26 May 2026
Safety

temporarily putting PhD in my social media name because this tweet from Elon is so asinine and because Elon’s goons went after a friend for …

DGX agent

temporarily putting PhD in my social media name because this tweet from Elon is so asinine and because Elon’s goons went after a friend for supporting the poor kid that Elon so rudely attacked. @iScie

safetygary-marcus--x
26 May 2026
Research

Terrain-Adaptive Grouser Wheel for Optimal Planetary Exploration: Design and Experimental Investigation

DGX agent

arXiv:2605.24311v1 Announce Type: new Abstract: Planetary rovers operating in extraterrestrial environments often encounter significant mobility challenges due to varying terrain features such as grad

researcharxiv-cs-ro
26 May 2026
Agents

Test-Time Deep Thinking to Explore Implicit Rules

DGX agent

arXiv:2605.24828v1 Announce Type: new Abstract: With the continuous advancement of Large Language Models (LLMs), intelligent agents are becoming increasingly vital. However, these agents often fail in

agentsarxiv-cs-ai
26 May 2026
Model Releases

Test-Time Graph Search for Goal-Conditioned Reinforcement Learning

DGX agent

arXiv:2510.07257v2 Announce Type: replace Abstract: Offline goal-conditioned reinforcement learning (GCRL) often struggles with long-horizon tasks, where errors in value estimation accumulate and prod

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

Test-Time Self-Adaptive Conditioning for Stable Audio-Driven Talking-Head Generation

DGX agent

arXiv:2605.25488v1 Announce Type: cross Abstract: Audio-driven talking-head generation has achieved remarkable progress with recent models such as AniTalker, FLOAT, and Sonic. Despite their success, m

model-releasesarxiv-cs-ai
26 May 2026
Research

Testing the Deliteralization Hypothesis in Human and Machine Translation

DGX agent

arXiv:2605.25686v1 Announce Type: new Abstract: The recent shift from dedicated NMT systems to general-purpose LLMs has reshaped machine translation, with LLMs reported to produce more fluent, less li

researcharxiv-cs-cl
26 May 2026
Research

TGFormer: Towards Temporal Graph Transformer with Auto-Correlation Mechanism

DGX agent

arXiv:2605.24971v1 Announce Type: cross Abstract: The growing interest in Temporal Graph Neural Networks (TGNNs) stems from their ability to model complex dynamics and deliver superior performance. Ho

researcharxiv-cs-ai
26 May 2026
Research

Thaka at KSAA-2026 Task 2: Regularized Fine-Tuning for Arabic Speech Diacritization

DGX agent

arXiv:2605.25928v1 Announce Type: new Abstract: We describe the winning system for Task 2 of the KSAA-2026 Shared Task on Arabic Speech Dictation with Automatic Diacritization. The task requires produ

researcharxiv-cs-cl
26 May 2026
Industry

🙏 Thank you all for the incredible love and support! Our latest Tencent Hunyuan translation models are on fire on Hugging Face: 🥰Hy-MT2-1.…

DGX agent

🙏 Thank you all for the incredible love and support! Our latest Tencent Hunyuan translation models are on fire on Hugging Face: 🥰Hy-MT2-1.8B ranks #1 🥰Hy-MT2-30B-A3B ranks #4 on the open-source model

industryclem-delangue--x
26 May 2026
Industry

Thank you so much for all the feedback on the Grok Build Beta. Some of you reported hitting limits quickly. Our team found areas to improve …

DGX agent

Thank you so much for all the feedback on the Grok Build Beta. Some of you reported hitting limits quickly. Our team found areas to improve caching, so we've reset Grok Build usage limits for all acco

industryelon-musk--x
26 May 2026
Industry

Thanks @AmericanAir for choosing @Starlink!

DGX agent

Elon Musk announced that American Airlines selected Starlink for connectivity services, highlighting a partnership between SpaceX's satellite internet provider and a major commercial airline. This rep

industryelon-musk--x
26 May 2026
Model Releases

The Age of Curiosity Meets the Age of AI: Benchmarking Child Safety in Large Language Models

DGX agent

arXiv:2605.25510v1 Announce Type: new Abstract: Children increasingly have access to Large Language Models (LLMs), which may expose them to responses that are developmentally inappropriate or require

model-releasesarxiv-cs-cl
26 May 2026
Model Releases

the basic trick to using Claude Code for non-technical work is to put a bunch of files in a folder and tell it can write scripts + make HTML

DGX agent

Claude Code can be used for non-technical work by organizing files in a folder and instructing it to write scripts and create HTML documents. This approach allows users without programming expertise t

model-releasesthariq--x
26 May 2026
Safety

The Behavioral Credibility Trilemma: When Calibrated Autonomy Becomes Impossible

DGX agent

arXiv:2605.25739v1 Announce Type: new Abstract: We prove that no reinforcement learning policy with confidence-gated autonomy can simultaneously achieve maximum helpfulness, optimal calibration, and f

safetyarxiv-cs-lg
26 May 2026
← Previous
1…10851086108710881089…1895
Next →