AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
Human
91,060Total entries
1Added by human
91,059Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
91,059 results
9 Jun 2026

ridiculous oversimplification, from a prominent OpenAI employee no less. does this capture your impression of how the two companies have beh…

SafetyDGX agent

ridiculous oversimplification, from a prominent OpenAI employee no less. does this capture your impression of how the two companies have behaved? The OAI / Anthropic values difference is deeply misund

RiskNet: A large-scale dataset of AI risk incidents from news with alignment and multi-dimensional annotations

Model ReleasesDGX agent

arXiv:2606.08376v1 Announce Type: cross Abstract: As artificial intelligence (AI) systems are increasingly deployed across socially consequential domains, reports of AI-related harms and failures have

RLVE: Scaling Up Reinforcement Learning for Language Models with Adaptive Verifiable Environments

Safety
DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2511.07317v2 Announce Type: replace-cross Abstract: We introduce Reinforcement Learning (RL) with Adaptive Verifiable Environments (RLVE), an approach using verifiable environments that procedur

Robot-DIFT: Correspondence-Sensitive Diffusion Features for Contact-Rich Robot Manipulation

SafetyDGX agent

arXiv:2602.11934v2 Announce Type: replace Abstract: Robot manipulation often fails in the final millimeters: a policy may recognize the right object yet miss the pose offsets, boundaries, or pre-conta

Robust In-Context Reinforcement Learning Under Reward Poisoning Attacks

ResearchDGX agent

arXiv:2506.06891v3 Announce Type: replace Abstract: We study the corruption-robustness of in-context reinforcement learning (ICRL), focusing on the Decision-Pretrained Transformer (DPT, Lee et al., 20

Robust Random Graph Matching in Dense Graphs via an Approximate Message Passing Type Algorithm

ResearchDGX agent

arXiv:2412.16457v3 Announce Type: replace-cross Abstract: In this paper, we focus on the matching recovery problem between a pair of correlated Gaussian Wigner matrices with a latent vertex correspond

Robust Renal Mass Segmentation on CT: A Validation Study of an AI-Based Framework

ResearchDGX agent

arXiv:2505.07573v2 Announce Type: replace-cross Abstract: Renal mass segmentation has important potential to enhance the clinical workflow, especially in settings requiring quantitative assessments. K

Robust-U1: Can MLLMs Self-Recover Corrupted Visual Content for Robust Understanding?

Model ReleasesDGX agent

arXiv:2606.08063v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have demonstrated remarkable success in visual understanding, yet their performance degrades significantly un

Rosetta Memory: Adaptive Memory for Cross-LLM Agents

Model ReleasesDGX agent

arXiv:2606.07711v1 Announce Type: cross Abstract: Memory is the key component for transforming a stateless LLM into a persistent, evolving agent through experience accumulation, long-horizon planning,

Routine laboratory trajectories encode the onset of organ-level complications in cancer

ApplicationsDGX agent

arXiv:2606.08538v1 Announce Type: new Abstract: Routine laboratory panels drawn during cancer treatment constitute longitudinal physiological recordings of organ function, yet their temporal structure

RPO-PDT: Demonstrating Role-Play-Based Knowledge Adaptation for Student Support Dialogue (Demonstration System)

SafetyDGX agent

arXiv:2606.09255v1 Announce Type: new Abstract: We present RPO-PDT: a retrieval-grounded, role-play-based dialogue system for adaptive student support in higher education. RPO-PDT is: (1) able to prov

RT-SDGOD: Real-Time Single-Domain Generalized Object Detection

ApplicationsDGX agent

arXiv:2606.09367v1 Announce Type: new Abstract: In real-world deployment under strict real-time constraints, weather and imaging variations induce significant distribution shifts, severely degrading d

RTL-BenchLS: A Large-Scale Benchmark for RTL Reasoning and Generation with Large Language Models

Model ReleasesDGX agent

arXiv:2606.08976v1 Announce Type: new Abstract: LLM-based RTL generation and reasoning is a promising direction for hardware design automation. High-quality benchmarks are critical infrastructure for

Rubrik turns its platform into an AI agent and ships Agent Cloud for Claude

Model ReleasesDGX agent

Rubrik Inc. today turned its data security platform into an autonomous agent and made its control layer for Anthropic PBC’s Claude generally available, the headline items in a wave of announcements at

Rule-based autocorrection of Piping and Instrumentation Diagrams (P&IDs) on graphs

ApplicationsDGX agent

arXiv:2502.18493v2 Announce Type: replace-cross Abstract: A piping and instrumentation diagram (P&ID) is a central reference document in chemical process engineering. Currently, chemical engineers man

RunAgent SuperBrowser: A Theory of Autonomous Web Navigation Grounded in Human Browsing Behaviour

Model ReleasesDGX agent

arXiv:2606.09399v1 Announce Type: new Abstract: We present SUPERBROWSER, an autonomous web-navigation agent designed against a single guiding hypothesis: a web agent should browse the way a person bro

SAD-Flower: Flow Matching for Safe, Admissible, and Dynamically Consistent Planning

SafetyDGX agent

arXiv:2511.05355v3 Announce Type: replace Abstract: Flow matching (FM) has shown promising results in data-driven planning. However, it inherently lacks formal guarantees for ensuring state and action

SAEExplainer: Interpreting SAE Features with Activation-Guided Preference Optimization

ResearchDGX agent

arXiv:2606.08496v1 Announce Type: cross Abstract: Although Sparse Autoencoders (SAEs) have mitigated the opacity of large language models (LLMs) by decomposing dense representations into sparse featur

Safe, Fluent and Acceptable Motion Generation and Execution for Human--Robot Interaction in Manufacturing Environments

SafetyDGX agent

arXiv:2606.08741v1 Announce Type: new Abstract: Robots operating in human environments must not only ensure physical safety but also exhibit behaviors that are understandable, fluent, and acceptable t

Safe Polytope-in-Polytope Motion Planning and Control with Control Barrier Functions

Local AiDGX agent

arXiv:2606.09719v1 Announce Type: new Abstract: Autonomous mobile robots operating in tight environments require motion planning frameworks that account for the physical footprint of the robot. Simpli

Safe-RULE: Safe Reinforcement UnLEarning

Model ReleasesDGX agent

arXiv:2606.09559v1 Announce Type: cross Abstract: Offline safe reinforcement learning (Safe RL) enables policy learning without online interactions, making it suitable for safety-critical systems such

SafeECGMatch: Calibration-Aware Joint Frequency and Time Space Semi-Supervised Learning for Open-Set ECG Classification

ResearchDGX agent

arXiv:2606.08037v1 Announce Type: cross Abstract: Electrocardiogram (ECG) classification models often suffer from severe label scarcity, making semi-supervised learning (SSL) an attractive strategy fo

SafeRun: Enabling Determinism in LLM Planning for Running

Model ReleasesDGX agent

arXiv:2606.09027v1 Announce Type: cross Abstract: Large Language Models enable flexible natural-language planning but remain unreliable in determinism-critical domains due to their probabilistic natur

Safety is Contextual, LLM-Judges Are Not: Navigating the Rigid Priors of Evaluators

SafetyDGX agent

arXiv:2606.07874v1 Announce Type: new Abstract: LLMs-as-judges are the only way to evaluate safety at scale. Despite their importance, LLM-judges themselves are rarely evaluated beyond human agreement

SAGE: An LLM-driven Self Reflective Agentic Framework for Fraud Detection

AgentsDGX agent

arXiv:2606.08146v1 Announce Type: new Abstract: Fraud detection in payment, e-commerce, and telecommunications systems requires accuracy at the individual level, robustness under severe class imbalanc

SAGE: Shape-Adapting Gated Experts for Adaptive Histopathology Image Segmentation

ResearchDGX agent

arXiv:2511.18493v4 Announce Type: replace-cross Abstract: The significant variability in cell size and shape continues to pose a major obstacle in computer-assisted cancer detection on gigapixel Whole

SailPoint shares fall despite earnings beat and raised guidance

IndustryDGX agent

Shares in SailPoint Inc. fell more than 11% today after the identity security company beat analyst expectations on revenue and adjusted earnings and raised its outlook, in a selloff that pointed to in

SAILS: Surrogate-based Analysis of Interactions via Local Effect Smooths

Local AiDGX agent

arXiv:2606.09404v1 Announce Type: cross Abstract: Feature interactions drive much of the predictive power of machine learning models, yet existing explanation methods only detect and quantify interact

Sample-Efficient LLM-Based Detection of Malicious Web Server Logs with Forensically Explainable Reasoning

TutorialsDGX agent

arXiv:2606.08649v1 Announce Type: cross Abstract: Forensic analysis of web server logs demands both accurate detection and human-readable explanations that can satisfy legal requirements. We present C

Sample-Efficient Post-Training for LEGO Spatial-Physics Reasoning

SafetyDGX agent

arXiv:2606.07602v1 Announce Type: cross Abstract: LLM-based LEGO assembly generation requires both semantic grounding and physical feasibility. We identify a data-induced failure mode, PhysHack, in wh

SAW: Stage-Aware Dynamic Weighting for Multi-Objective Reinforcement Learning in Large Language Models

SafetyDGX agent

arXiv:2606.07705v1 Announce Type: cross Abstract: Although multi-objective reinforcement learning (MORL) is central to aligning large language models with complex human preferences, the prevailing pra

SC3: The Multi-Solvent Solubility Challenge and Benchmark

Model ReleasesDGX agent

arXiv:2606.07656v1 Announce Type: cross Abstract: Solubility prediction is a standard benchmark in computational chemistry, yet multi-solvent models which reportedly approach the experimental-noise ce

Scaffold Effects on GAIA: A Controlled Comparison

Model ReleasesDGX agent

arXiv:2606.08529v1 Announce Type: new Abstract: Published agent capability scores conflate what a model can do with what its scaffold lets it do, and the magnitude of this elicitation gap is not well

Scale Robot Reinforcement Learning with NVIDIA Isaac Lab on Amazon SageMaker AI

HardwareDGX agent

In this post, we show how to train robot policies for the Unitree H1 humanoid with NVIDIA Isaac Lab on Amazon SageMaker AI across two compute options: Amazon SageMaker HyperPod and Amazon SageMaker Tr

ScaleSweep: Accurate NVFP4 Post-Training Quantization of LLMs via Block Scale Initialization

Model ReleasesDGX agent

arXiv:2606.07618v1 Announce Type: cross Abstract: NVFP4 is a recently introduced hardware-supported FP4 format that improves the fidelity of 4-bit quantization through fine-grained block scales. Howev

Scaling by Diversified Experience for Vision-Language-Action Models

Model ReleasesDGX agent

arXiv:2606.09009v1 Announce Type: new Abstract: Vision-Language-Action models face significant challenges in real-world deployment due to the entanglement of high-level reasoning with low-level contro

Scaling Decision-Focused Learning to Large Problems with Lagrangian Decomposition

ResearchDGX agent

arXiv:2606.08797v1 Announce Type: cross Abstract: Decision-focused learning has shown great promise for addressing predict-then-optimize problems, particularly in the presence of under-specified model

Scaling Laws for Masked-Reconstruction Transformers on Single-Cell Transcriptomics

Model ReleasesDGX agent

arXiv:2602.15253v2 Announce Type: replace Abstract: Neural scaling laws -- power-law relationships between loss, model size, and data -- have been extensively documented for language and vision transf

Scaling Neural Network Verification with Tensor Parallelism and Fully Sharded Data Parallelism

SafetyDGX agent

arXiv:2606.09377v1 Announce Type: cross Abstract: Formal neural network verification -- proving that a network satisfies safety properties for all inputs in a specified domain -- is bounded in practic

Scaling Participation in Modular AI Systems

ResearchDGX agent

arXiv:2606.07812v1 Announce Type: new Abstract: Humanity is a mosaic of multifaceted talents and needs, and any truly intelligent AI must reflect that richness. Yet the LLMs used by all are built by t

scCBGM: Interpretable Single-Cell Counterfactual Editing

Model ReleasesDGX agent

arXiv:2606.07760v1 Announce Type: new Abstract: Understanding cellular phenotypes and how they respond to perturbations is critical for disease biology and therapeutic design. Single-cell RNA sequenci

SceneConductor: 3D Scene Generation from Single Image with Multi-Agent Orchestration

Model ReleasesDGX agent

arXiv:2606.08402v1 Announce Type: cross Abstract: Generating complete 3D scenes from a single image requires inferring globally consistent geometry, object relationships, and environmental context fro

Sci-Rho: A Multilingual Visually-Grounded Symbolic Benchmark for STEM Problems

Model ReleasesDGX agent

arXiv:2606.08034v1 Announce Type: cross Abstract: Symbolic benchmarks have emerged as a key approach to assess model robustness under minor modifications to STEM-related questions. However, existing s

SciFlow-Bench: Evaluating Structure-Aware Scientific Diagram Generation via Inverse Parsing

Model ReleasesDGX agent

arXiv:2602.09809v2 Announce Type: replace Abstract: Scientific diagrams convey explicit structural information, yet modern text-to-image models often produce visually plausible but structurally incorr

SciTrace: Trajectory-Aware Safety Reasoning for Scientific Discovery Agents

SafetyDGX agent

arXiv:2606.08234v1 Announce Type: new Abstract: LLM-based scientific agents have shown strong capacity for autonomous research, yet their safety layers remain structurally divorced from core reasoning

Screwworms in US: Human risk is low—but they can burrow through your skull

IndustryDGX agent

New World screwworms were detected in the United States in June 2026 for the first time in decades , with no reported human infestations and low human risk . When humans are infected, infestations cau

SDTrack: A Baseline for Event-based Tracking via Spiking Neural Networks

ResearchDGX agent

arXiv:2503.08703v4 Announce Type: replace-cross Abstract: Event cameras provide superior temporal resolution, dynamic range, energy efficiency, and pixel bandwidth. Spiking Neural Networks (SNNs) natu

SearchSwarm: Towards Delegation Intelligence in Agentic LLMs for Long-Horizon Deep Research

AgentsDGX agent

arXiv:2606.09730v1 Announce Type: new Abstract: Large language models are increasingly expected to handle complex, long-horizon real-world tasks whose context demands can grow without bound, yet model

Seattle City Council votes 9-0 to enact a one-year moratorium on new large data centers and study their impact; Mayor Katie Wilson is expected to sign the bill (Greg Kim/The Seattle Times)

IndustryDGX agent

Greg Kim / The Seattle Times: Seattle City Council votes 9-0 to enact a one-year moratorium on new large data centers and study their impact; Mayor Katie Wilson is expected to sign the bill — Amid a g

SecureClaw: Clawing Back Control of LLM Agents

SafetyDGX agent

arXiv:2606.09549v1 Announce Type: cross Abstract: Tool-using large language model (LLM) agents face two distinct security failures: unauthorized external actions and exposure of sensitive plaintext in

Securing Self-supervised Data Curation for Foundation Models Robustness

ResearchDGX agent

arXiv:2606.09511v1 Announce Type: new Abstract: Self-supervised data curation provides a pathway to scaling and improving the generalization capabilities of machine learning models. By leveraging self

See how Claude Fable 5 compares across every model: http://cursor.com/evals

Model ReleasesDGX agent

Claude Fable 5 is compared against other AI models on various evaluation metrics through Cursor's benchmarking tool. The evaluation likely covers performance across different tasks such as coding, rea

See More, Match Better: Multi-Source Feature Fusion for Two-View Correspondence Learning

SafetyDGX agent

arXiv:2606.09262v1 Announce Type: new Abstract: Two-view correspondence learning aims to distinguish true correspondences (inliers) from false ones (outliers) in image pairs by leveraging their underl

See More, Think Deeper: Query-Expanded Visual Evidence and Answer-Clue Guided Reflection for Long Video Understanding

Model ReleasesDGX agent

arXiv:2606.09064v1 Announce Type: cross Abstract: Recent advances in Video Large Language Models (Video-LLMs) have enabled performance on long-video understanding tasks. However, existing methods stil

Seeing is Believing: Aligning Prompt Rewriting with Visual Anchors for Text-to-Image Generation

ResearchDGX agent

arXiv:2606.08492v1 Announce Type: cross Abstract: Despite the impressive capabilities of text-to-image (T2I) models, an intent-generation gap often persists due to the brevity and ambiguity of user pr

Seeing the Hivemind: A Consensus-Aware Interaction Technique for Mitigating AI Homogenization

Local AiDGX agent

arXiv:2606.09587v1 Announce Type: cross Abstract: People are increasingly using AI for creative tasks such as writing. While adoption continues to grow, this form of use risks undermining individual c

SEF-CLGC at SemEval-2026 Task 11: Logical Notation Impact on Language Model Performance

SafetyDGX agent

arXiv:2606.09157v1 Announce Type: cross Abstract: This paper revisits our pipeline called Syllogistic Evaluation Framework-Common Logic Grammar Construction (SEF-CLGC). We combine formal logical notat

Segment-level Tree Search for Long Meeting Document Summarization

ResearchDGX agent

arXiv:2606.08445v1 Announce Type: cross Abstract: Meeting documents are challenging to summarize due to their length and complex conversational structure. Existing approaches typically adopt multi-sta

SegmentAnyTreeV2: Scaling Transformer-Based Tree Instance Segmentation Across Sensors, Platforms, and Forests

Model ReleasesDGX agent

arXiv:2606.08206v1 Announce Type: new Abstract: We present SegmentAnyTreeV2, a sensor- and platform-agnostic framework for semantic and instance segmentation of forest point clouds. The model combines

Segmentation-Assisted Brain MRI Synthesis with Cross-Image Multi-Contrast Feature Memory Bank Retrieval Augmentation

ResearchDGX agent

arXiv:2606.08421v1 Announce Type: new Abstract: Multi-contrast brain MRI provide complementary soft-tissue characteristics that aid in the screening and diagnosis of diseases. However, limited scannin

← Previous
1…652653654655656…1518
Next →