AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,814
  • Agents7,519
  • Applications5,378
  • Concepts5
  • Hardware1,822
  • Industry6,162
  • Local Ai4,908
  • Model Releases23,658
  • Research20,008
  • Safety13,291
  • Syntheses17
  • Tools1,674
  • Tutorials3,372

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,814
  • Agents7,519
  • Applications5,378
  • Concepts5
  • Hardware1,822
  • Industry6,162
  • Local Ai4,908
  • Model Releases23,658
  • Research20,008
  • Safety13,291
  • Syntheses17
  • Tools1,674
  • Tutorials3,372

Source
HumanDGX agent

Content type
87,814Total entries
1Added by human
87,813Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
87,813 results
Safety

Repeated Deceptive Path Planning against Learnable Observer

DGX agent

arXiv:2605.07174v1 Announce Type: new Abstract: We study the problem of deceptive path planning (DPP), where an agent aims to conceal its true destination from external observers. While existing work

safetyarxiv-cs-ai
11 May 2026
Research

Replicating Human Motivated Reasoning Studies with LLMs

AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2601.16130v2 Announce Type: replace-cross Abstract: Motivated reasoning - the idea that individuals processing information may be motivated to either arrive at accurate beliefs or arrive at desi

researcharxiv-cs-ai
11 May 2026
Safety

reply to Hinton’s reply to me, for additional context:

DGX agent

reply to Hinton’s reply to me, for additional context: Dear @geoffreyhinton, I literally never said that AI systems “JUST regurgitate”; that’s plainly false. I don’t believe it, and I didn’t say it. (

safetygary-marcus--x
11 May 2026
Industry

Report: AI chipmaker Cerebras to increase IPO price target amid surging investor demand

DGX agent

Artificial intelligence chipmaker Cerebras Systems Inc. is expected to increase the size and price of its initial public offering later today as investor demand for access to its shares continues to r

industrysiliconangle
11 May 2026
Model Releases

ReSeek: A Self-Correcting Framework for Search Agents with Instructive Rewards

DGX agent

arXiv:2510.00568v3 Announce Type: replace Abstract: Search agents powered by Large Language Models (LLMs) have demonstrated significant potential in tackling knowledge-intensive tasks. Reinforcement l

model-releasesarxiv-cs-cl
11 May 2026
Safety

Resource-Element Energy Difference for Noncoherent Over-the-Air Federated Learning

DGX agent

arXiv:2605.07263v1 Announce Type: cross Abstract: Over-the-air federated learning (OTA-FL) reduces uplink latency by exploiting waveform superposition, but conventional analog aggregation schemes typi

safetyarxiv-cs-ai
11 May 2026
Safety

Response-G1: Explicit Scene Graph Modeling for Proactive Streaming Video Understanding

DGX agent

arXiv:2605.07575v1 Announce Type: cross Abstract: Proactive streaming video understanding requires Video-LLMs to decide when to respond as a video unfolds, a task where existing methods often fall sho

safetyarxiv-cs-ai
11 May 2026
Safety

Response Time Enhances Alignment with Heterogeneous Preferences

DGX agent

arXiv:2605.06987v1 Announce Type: new Abstract: Aligning large language models (LLMs) to human preferences typically relies on aggregating pooled feedback into a single reward model. However, this sta

safetyarxiv-cs-lg
11 May 2026
Industry

Retail markdown optimization: from reactive markdowns to proactive

DGX agent

This Databricks resource examines retail markdown optimization strategies, contrasting reactive markdown approaches (responding to inventory issues after they occur) with proactive markdown strategies

industrydatabricks
11 May 2026
Model Releases

Rethinking Dense Optical Flow without Test-Time Scaling

DGX agent

arXiv:2605.08000v1 Announce Type: new Abstract: Recent progress in dense optical flow has been driven by increasingly complex architectures and multi-step refinement for test-time scaling. While these

model-releasesarxiv-cs-cv
11 May 2026
Research

Rethinking Dense Sequential Chains: Reasoning Language Models Can Extract Answers from Sparse, Order-Shuffling Chain-of-Thoughts

DGX agent

arXiv:2605.07307v1 Announce Type: new Abstract: Modern reasoning language models generate dense, sequential chain-of-thought traces implicitly assuming that every token contributes and that steps must

researcharxiv-cs-cl
11 May 2026
Research

Rethinking Experience Utilization in Self-Evolving Language Model Agents

DGX agent

arXiv:2605.07164v1 Announce Type: new Abstract: Self-evolving agents improve by accumulating and reusing experience from past interactions. Existing work has largely focused on how experience is const

researcharxiv-cs-cl
11 May 2026
Safety

Rethinking Importance Sampling in LLM Policy Optimization: A Cumulative Token Perspective

DGX agent

arXiv:2605.07331v1 Announce Type: cross Abstract: Reinforcement learning, including reinforcement learning with verifiable rewards (RLVR), has emerged as a powerful approach for LLM post-training. Cen

safetyarxiv-cs-ai
11 May 2026
Tutorials

Rethinking State Tracking in Recurrent Models Through Error Control Dynamics

DGX agent

arXiv:2605.07755v1 Announce Type: cross Abstract: The theory of state tracking in recurrent architectures has predominantly focused on expressive capacity: whether a fixed architecture can theoretical

tutorialsarxiv-cs-cl
11 May 2026
Model Releases

Rethinking Weight Tying: Pseudo-Inverse Tying for LM Stable Training and Updates

DGX agent

arXiv:2602.04556v2 Announce Type: replace Abstract: Weight tying is widely used in compact language models to reduce parameters by sharing the token table between the input embedding and the output pr

model-releasesarxiv-cs-cl
11 May 2026
Model Releases

Retina-RAG: Retrieval-Augmented Vision-Language Modeling for Joint Retinal Diagnosis and Clinical Report Generation

DGX agent

arXiv:2605.06173v2 Announce Type: replace-cross Abstract: Diabetic Retinopathy (DR) is a leading cause of preventable blindness among working-age adults worldwide, yet most automated screening systems

model-releasesarxiv-cs-ai
11 May 2026
Research

Retrieval from Within: An Intrinsic Capability of Attention-Based Models

DGX agent

arXiv:2605.05806v2 Announce Type: replace Abstract: Retrieval-augmented generation (RAG) typically treats retrieval and generation as separate systems. We ask whether an attention-based encoder-decode

researcharxiv-cs-lg
11 May 2026
Research

Retrieval Heads are Dynamic

DGX agent

arXiv:2602.11162v2 Announce Type: replace Abstract: Recent studies have identified 'retrieval heads' in Large Language Models (LLMs) responsible for extracting information from input contexts. However

researcharxiv-cs-cl
11 May 2026
Research

Retrieve, Integrate, and Synthesize: Spatial-Semantic Grounded Latent Visual Reasoning

DGX agent

arXiv:2605.07106v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have made remarkable progress on vision-language reasoning, yet most methods still compress visual evidence int

researcharxiv-cs-cl
11 May 2026
Research

Revisiting Adam for Streaming Reinforcement Learning

DGX agent

arXiv:2605.06764v1 Announce Type: cross Abstract: Learning from a sequence of interactions, as soon as observations are perceived and acted upon, without explicitly storing them, holds the promise of

researcharxiv-cs-ai
11 May 2026
Model Releases

Revisiting Transformer Layer Parameterization Through Causal Energy Minimization

DGX agent

arXiv:2605.07588v1 Announce Type: cross Abstract: Transformer blocks typically combine multi-head attention (MHA) for token mixing with gated MLPs for token-wise feature transformation, yet many choic

model-releasesarxiv-cs-ai
11 May 2026
Safety

RIDER: 3D RNA Inverse Design with Reinforcement Learning-Guided Diffusion

DGX agent

arXiv:2602.16548v2 Announce Type: replace Abstract: The inverse design of RNA three-dimensional (3D) structures is crucial for engineering functional RNAs in synthetic biology and therapeutics. While

safetyarxiv-cs-lg
11 May 2026
Safety

Risk-Consistent Multiclass Learning from Random Label-Subset Membership Queries

DGX agent

arXiv:2605.07413v1 Announce Type: new Abstract: Obtaining accurate class labels is often costly or unreliable, and may also be limited by privacy or other practical conditions. Compared with asking an

safetyarxiv-cs-lg
11 May 2026
Research

RL-RIG: A Generative Spatial Reasoner via Intrinsic Reflection

DGX agent

arXiv:2602.19974v2 Announce Type: replace Abstract: Recent advancements in image generation have achieved impressive results in producing high-quality images. However, existing image generation models

researcharxiv-cs-cv
11 May 2026
Local Ai

RNAGenScape: Property-Guided, Optimized Generation of mRNA Sequences with Manifold Langevin Dynamics

DGX agent

arXiv:2510.24736v3 Announce Type: replace-cross Abstract: Generating property-optimized mRNA sequences is central to applications such as vaccine design and protein replacement therapy, but remains ch

local-aiarxiv-cs-lg
11 May 2026
Applications

Robust and Reliable AI for Predictive Quality in Semiconductor Materials Manufacturing with MLOps and Uncertainty Quantification

DGX agent

arXiv:2605.07752v1 Announce Type: new Abstract: Semiconductor materials manufacturing presents unique challenges for machine learning deployment due to evolving process conditions, equipment degradati

applicationsarxiv-cs-lg
11 May 2026
Research

Robust stochastic first order methods in heavy-tailed noise via medoid mini-batch gradient sampling

DGX agent

arXiv:2605.07634v1 Announce Type: cross Abstract: We consider a first order stochastic optimization framework where, at each iteration, K independent identically distributed (i.i.d.) data point sample

researcharxiv-cs-lg
11 May 2026
Model Releases

Robust Sublinear Convergence Rates for Iterative Bregman Projections

DGX agent

arXiv:2602.01372v2 Announce Type: replace-cross Abstract: Entropic regularization provides a simple way to approximate linear programs whose constraints split into two or more tractable blocks. The re

model-releasesarxiv-cs-lg
11 May 2026
Safety

Robustness of Refugee-Matching Gains to Off-Policy Evaluation Choices

DGX agent

arXiv:2605.06686v1 Announce Type: new Abstract: Previous research has investigated the potential of refugee matching for boosting refugee outcomes, first considered by Bansak et al. (2018). This paper

safetyarxiv-cs-lg
11 May 2026
Safety

Rollback-Free Stable Brick Structures Generation

DGX agent

arXiv:2605.06947v1 Announce Type: new Abstract: While autoregressive models have advanced 3D generation, creating physically stable brick structures remains a challenge due to the strict requirements

safetyarxiv-cs-lg
11 May 2026
Model Releases

RRCM: Ranking-Driven Retrieval over Collaborative and Meta Memories for LLM Recommendation

DGX agent

arXiv:2605.07129v1 Announce Type: cross Abstract: Large Language Models (LLMs) have emerged as a promising paradigm for next-generation recommender systems, offering strong semantic understanding and

model-releasesarxiv-cs-ai
11 May 2026
Safety

Rubric-based On-policy Distillation

DGX agent

arXiv:2605.07396v1 Announce Type: cross Abstract: On-policy distillation (OPD) is a powerful paradigm for model alignment, yet its reliance on teacher logits restricts its application to white-box sce

safetyarxiv-cs-ai
11 May 2026
Model Releases

Rubric-Grounded RL: Structured Judge Rewards for Generalizable Reasoning

DGX agent

arXiv:2605.08061v1 Announce Type: new Abstract: We argue that decomposing reward into weighted, verifiable criteria and using an LLM judge to score them provides a partial-credit optimization signal:

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

RuleSafe-VL: Evaluating Rule-Conditioned Decision Reasoning in Vision-Language Content Moderation

DGX agent

arXiv:2605.07760v1 Announce Type: new Abstract: Platform content moderation applies explicit policy rules and context-dependent conditions to decide whether user content is allowed, restricted, or rem

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

S2M-Net: Spectral-Spatial Mixing for Medical Image Segmentation with Morphology-Aware Adaptive Loss

DGX agent

arXiv:2601.01285v2 Announce Type: replace Abstract: Medical image segmentation requires balancing local precision for boundary-critical clinical applications, global context for anatomical coherence,

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

S2S-Arena: Evaluating Paralinguistic Instruction Following in Speech-to-Speech Models

DGX agent

arXiv:2503.05085v2 Announce Type: replace Abstract: Recent advances in large language models (LLMs) have fundamentally reshaped speech-to-speech (S2S) systems, enabling increasingly natural spoken int

model-releasesarxiv-cs-cl
11 May 2026
Safety

Safactory: A Scalable Agentic Infrastructure for Training Trustworthy Autonomous Intelligence

DGX agent

arXiv:2605.06230v2 Announce Type: replace Abstract: As large models evolve from conversational assistants into autonomous agents, challenges increasingly arise from long-horizon decision making, tool

safetyarxiv-cs-ai
11 May 2026
Model Releases

Safe, or Simply Incapable? Rethinking Safety Evaluation for Phone-Use Agents

DGX agent

arXiv:2605.07630v1 Announce Type: cross Abstract: When a phone-use agent avoids harm, does that show safety, or simply inability to act? Existing evaluations often cannot tell. A harmful outcome may b

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Safety Anchor: Defending Harmful Fine-tuning via Geometric Bottlenecks

DGX agent

arXiv:2605.05995v2 Announce Type: replace-cross Abstract: The safety alignment of Large Language Models (LLMs) remains vulnerable to Harmful Fine-tuning (HFT). While existing defenses impose constrain

model-releasesarxiv-cs-ai
11 May 2026
Safety

SAGE: Hierarchical LLM-Based Literary Evaluation through Ontology-Grounded Interpretive Dimensions

DGX agent

arXiv:2605.07102v1 Announce Type: new Abstract: Evaluating literary quality requires assessing interpretive dimensions such as cultural representation, emotional depth, and philosophical sophisticatio

safetyarxiv-cs-cl
11 May 2026
Research

Saliency-Aware Regularized Quantization Calibration for Large Language Models

DGX agent

arXiv:2605.05693v2 Announce Type: replace Abstract: Post-training quantization (PTQ) is an effective approach for deploying large language models (LLMs) under memory and latency constraints. Most exis

researcharxiv-cs-ai
11 May 2026
Research

SAM 3D Animal: Promptable Animal 3D Reconstruction from Images in the Wild

DGX agent

arXiv:2605.07604v1 Announce Type: cross Abstract: 3D animal reconstruction in the wild remains challenging due to large species variation, frequent occlusions, and the prevalence of multi-animal scene

researcharxiv-cs-ai
11 May 2026
Tutorials

Same Brain, Different Prediction: How Preprocessing Choices Undermine EEG Decoding Reliability

DGX agent

arXiv:2605.07212v1 Announce Type: cross Abstract: Electroencephalography (EEG) is a cornerstone of brain-computer interfaces and clinical neuroscience, yet deep learning models are typically trained a

tutorialsarxiv-cs-ai
11 May 2026
Safety

Same Signal, Opposite Meaning: Direction-Informed Adaptive Learning for LLM Agents

DGX agent

arXiv:2605.06908v1 Announce Type: cross Abstract: Adaptive test-time compute for LLM agents aims to invoke extra computation only when it improves performance. Existing methods typically use confidenc

safetyarxiv-cs-ai
11 May 2026
Research

Sample Complexity of Stochastic Optimization with Integer Variables

DGX agent

arXiv:2605.07239v1 Announce Type: new Abstract: We establish sample complexity results for stochastic optimization over the integers, especially with a view to understand the complexity with respect t

researcharxiv-cs-lg
11 May 2026
Industry

Samsung made a “mockery” of Dua Lipa by putting her picture on TV boxes, lawsuit says

DGX agent

Dua Lipa filed a $15 million lawsuit against Samsung, alleging the electronics company used her photograph to sell televisions without permission or payment. Samsung began featuring Lipa's image on TV

industryars-technica
11 May 2026
Safety

SARA: Semantically Adaptive Relational Alignment for Video Diffusion Models

DGX agent

arXiv:2605.07800v1 Announce Type: new Abstract: Recent video diffusion models (VDMs) synthesize visually convincing clips, yet still drop entities, mis-bind attributes, and weaken the interactions spe

safetyarxiv-cs-cv
11 May 2026
Applications

Sarra builds apps with her two sons. Noni shipped Bamboo Brain into the App Store's top 12 in Education. Rebecca built the system she wished…

DGX agent

Sarra builds apps with her two sons. Noni shipped Bamboo Brain into the App Store's top 12 in Education. Rebecca built the system she wished she'd had during her custody battle. To every mother buildi

applicationsreplit--x
11 May 2026
← Previous
1…13201321132213231324…1830
Next →