AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
83,193 results
10 Apr 2026

Reason in Chains, Learn in Trees: Self-Rectification and Grafting for Multi-turn Agent Policy Optimization

SafetyDGX agent

arXiv:2604.07165v1 Announce Type: new Abstract: Reinforcement learning for Large Language Model agents is often hindered by sparse rewards in multi-step reasoning tasks. Existing approaches like Group

Reason-SVG: Enhancing Structured Reasoning for Vector Graphics Generation with Reinforcement Learning

SafetyDGX agent

arXiv:2505.24499v2 Announce Type: replace Abstract: Generating high-quality Scalable Vector Graphics (SVGs) is challenging for Large Language Models (LLMs), as it requires advanced reasoning for struc

Reasoning-Based Refinement of Unsupervised Text Clusters with LLMs

ApplicationsDGX agent

arXiv:2604.07562v1 Announce Type: new Abstract: Unsupervised methods are widely used to induce latent semantic structure from large text collections, yet their outputs often contain incoherent, redund

Content type
AllBlogX PostPaperYouTubeRedditGitHub

Reasoning Fails Where Step Flow Breaks

ResearchDGX agent

arXiv:2604.06695v1 Announce Type: new Abstract: Large reasoning models (LRMs) that generate long chains of thought now perform well on multi-step math, science, and coding tasks. However, their behavi

Reasoning Graphs: Deterministic Agent Accuracy through Evidence-Centric Chain-of-Thought Feedback

AgentsDGX agent

arXiv:2604.07595v1 Announce Type: cross Abstract: Language model agents reason from scratch on every query: each time an agent retrieves evidence and deliberates, the chain of thought is discarded and

Reasoning Within the Mind: Dynamic Multimodal Interleaving in Latent Space

SafetyDGX agent

arXiv:2512.12623v3 Announce Type: replace-cross Abstract: Recent advancements in Multimodal Large Language Models (MLLMs) have significantly enhanced cross-modal understanding and reasoning by incorpo

ReCellTy: Domain-Specific Knowledge Graph Retrieval-Augmented LLMs Reasoning Workflow for Single-Cell Annotation

ResearchDGX agent

arXiv:2505.00017v2 Announce Type: replace Abstract: With the rapid development of large language models (LLMs), their application to cell type annotation has drawn increasing attention. However, gener

Recommended Model for a 4060ti 8gb and 16gb ram

Local AiDGX agent

For users running Ollama on an NVIDIA RTX 4060 Ti with 8GB VRAM and 16GB system RAM, the community consensus recommends 7B–8B parameter models (such as Llama 3.1 8B, Mistral 7B, or Qwen 8B) using Q...

ReconPhys: Reconstruct Appearance and Physical Attributes from Single Video

ResearchDGX agent

arXiv:2604.07882v1 Announce Type: new Abstract: Reconstructing non-rigid objects with physical plausibility remains a significant challenge. Existing approaches leverage differentiable rendering for p

RectifiedHR: Enable Efficient High-Resolution Synthesis via Energy Rectification

ResearchDGX agent

arXiv:2503.02537v4 Announce Type: replace Abstract: Diffusion models have achieved remarkable progress across various visual generation tasks. However, their performance significantly declines when ge

Rectifying LLM Thought from Lens of Optimization

ResearchDGX agent

arXiv:2512.01925v2 Announce Type: replace-cross Abstract: Recent advancements in large language models (LLMs) have been driven by their emergent reasoning capabilities, particularly through long chain

ReDAct: Uncertainty-Aware Deferral for LLM Agents

AgentsDGX agent

arXiv:2604.07036v1 Announce Type: cross Abstract: Recently, LLM-based agents have become increasingly popular across many applications, including complex sequential decision-making problems. However,

Reflection-Based Task Adaptation for Self-Improving VLA

SafetyDGX agent

arXiv:2510.12710v3 Announce Type: replace Abstract: Pre-trained Vision-Language-Action (VLA) models represent a major leap towards general-purpose robots, yet efficiently adapting them to novel, speci

ReflectRM: Boosting Generative Reward Models via Self-Reflection within a Unified Judgment Framework

SafetyDGX agent

arXiv:2604.07506v1 Announce Type: cross Abstract: Reward Models (RMs) are critical components in the Reinforcement Learning from Human Feedback (RLHF) pipeline, directly determining the alignment qual

Region-Graph Optimal Transport Routing for Mixture-of-Experts Whole-Slide Image Classification

Local AiDGX agent

arXiv:2604.07298v1 Announce Type: cross Abstract: Multiple Instance Learning (MIL) is the dominant framework for gigapixel whole-slide image (WSI) classification in computational pathology. However, c

Region-R1: Reinforcing Query-Side Region Cropping for Multi-Modal Re-Ranking

SafetyDGX agent

arXiv:2604.05268v2 Announce Type: replace-cross Abstract: Multi-modal retrieval-augmented generation (MM-RAG) relies heavily on re-rankers to surface the most relevant evidence for image-question quer

Reinforcement-Guided Synthetic Data Generation for Privacy-Sensitive Identity Recognition

Model ReleasesDGX agent

arXiv:2604.07884v1 Announce Type: new Abstract: High-fidelity generative models are increasingly needed in privacy-sensitive scenarios, where access to data is severely restricted due to regulatory an

Relatively smart for a human, but not by AI standards

IndustryDGX agent

I was unable to retrieve the specific post at the URL provided. The status ID `2042488382438482407` does not appear in any accessible search results, and X (formerly Twitter) requires JavaScript an...

Remember when ChatGPT helped someone plan a mass shooting, and a dozen OpenAI employees notified management, and they did nothing and 8 peop…

SafetyDGX agent

Remember when ChatGPT helped someone plan a mass shooting, and a dozen OpenAI employees notified management, and they did nothing and 8 people died? They don't want to be liable for their incompetence

RemoteAgent: Bridging Vague Human Intents and Earth Observation with RL-based Agentic MLLMs

AgentsDGX agent

arXiv:2604.07765v1 Announce Type: new Abstract: Earth Observation (EO) systems are essentially designed to support domain experts who often express their requirements through vague natural language ra

Replacing Tunable Parameters in Weather and Climate Models with State-Dependent Functions using Reinforcement Learning

Model ReleasesDGX agent

arXiv:2601.04268v2 Announce Type: replace Abstract: Weather and climate models rely on parametrisations to represent unresolved sub-grid processes. Traditional schemes rely on fixed coefficients that

Replit now deploys directly to Databricks. Your apps run inside your Databricks environment while inheriting its security, governance, and d…

ToolsDGX agent

Replit now deploys directly to Databricks. Your apps run inside your Databricks environment while inheriting its security, governance, and data access. Beta is live. Enterprises are already building w

Report: US demands Reddit unmask ICE critic, summons firm to grand jury

IndustryDGX agent

The Trump administration has ordered Reddit to appear before a secret grand jury in Washington, D.C., demanding the personal data of an anonymous user who criticized ICE online, with a compliance d...

Research with ChatGPT

TutorialsDGX agent

The OpenAI Academy 'Research with ChatGPT' tutorial covers how to use ChatGPT's two primary research tools: the Search the Web feature for retrieving live data and sources quickly, and Deep Researc...

Reset-Free Reinforcement Learning for Real-World Agile Driving: An Empirical Study

SafetyDGX agent

arXiv:2604.07672v1 Announce Type: new Abstract: This paper presents an empirical study of reset-free reinforcement learning (RL) for real-world agile driving, in which a physical 1/10-scale vehicle le

Resistance Distance and Linearized Optimal Transport on Graphs

ResearchDGX agent

arXiv:2404.15261v4 Announce Type: replace-cross Abstract: We study the linearization of a discrete transportation distance between probability distributions on finite weighted graphs originally due to

Resource-constrained Amazons chess decision framework integrating large language models and graph attention

ResearchDGX agent

arXiv:2603.10512v2 Announce Type: replace Abstract: Artificial intelligence has advanced significantly through the development of intelligent game-playing systems, providing rigorous testbeds for deci

Responsible and safe use of AI

SafetyDGX agent

The OpenAI Academy page on 'Responsible and Safe Use of AI' is a guidance resource focused on best practices for using ChatGPT responsibly in professional and personal settings. It emphasizes that ...

Restoring Heterogeneity in LLM-based Social Simulation: An Audience Segmentation Approach

Model ReleasesDGX agent

arXiv:2604.06663v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used to simulate social attitudes and behaviors, offering scalable 'silicon samples' that can approximat

Rethinking Data Mixing from the Perspective of Large Language Models

ResearchDGX agent

arXiv:2604.07963v1 Announce Type: new Abstract: Data mixing strategy is essential for large language model (LLM) training. Empirical evidence shows that inappropriate strategies can significantly redu

Rethinking Entropy Allocation in LLM-based ASR: Understanding the Dynamics between Speech Encoders and LLMs

Model ReleasesDGX agent

arXiv:2604.08003v1 Announce Type: cross Abstract: Integrating large language models (LLMs) into automatic speech recognition (ASR) has become a dominant paradigm. Although recent LLM-based ASR models

Rethinking Generalization in Reasoning SFT: A Conditional Analysis on Optimization, Data, and Model Capability

SafetyDGX agent

arXiv:2604.06628v1 Announce Type: new Abstract: A prevailing narrative in LLM post-training holds that supervised finetuning (SFT) memorizes while reinforcement learning (RL) generalizes. We revisit t

REVEAL: Reasoning-Enhanced Forensic Evidence Analysis for Explainable AI-Generated Image Detection

Model ReleasesDGX agent

arXiv:2511.23158v2 Announce Type: replace-cross Abstract: The rapid progress of visual generative models has made AI-generated images increasingly difficult to distinguish from authentic ones, posing

Revisiting Fairness Impossibility with Endogenous Behavior

SafetyDGX agent

arXiv:2604.06378v1 Announce Type: cross Abstract: In many real-world settings, institutions can and do adjust the consequences attached to algorithmic classification decisions, such as the size of fin

Revisiting Radar Perception With Spectral Point Clouds

Model ReleasesDGX agent

arXiv:2604.08282v1 Announce Type: new Abstract: Radar perception models are trained with different inputs, from range-Doppler spectra to sparse point clouds. Dense spectra are assumed to outperform sp

RewardFlow: Generate Images by Optimizing What You Reward

SafetyDGX agent

arXiv:2604.08536v1 Announce Type: new Abstract: We introduce RewardFlow, an inversion-free framework that steers pretrained diffusion and flow-matching models at inference time through multi-reward La

Riemann-Bench: A Benchmark for Moonshot Mathematics

Model ReleasesDGX agent

arXiv:2604.06802v1 Announce Type: new Abstract: Recent AI systems have achieved gold-medal-level performance on the International Mathematical Olympiad, demonstrating remarkable proficiency at competi

RISC-V chip design startup SiFive nabs $400M investment

HardwareDGX agent

SiFive Inc., a startup that sells chip designs based on the open-source RISC-V architecture, has raised $400 million in funding at a $3.65 billion valuation. Atreides Management led the Series G round

RLBoost: Harvesting Preemptible Resources for Cost-Efficient Reinforcement Learning on LLMs

HardwareDGX agent

arXiv:2510.19225v3 Announce Type: replace-cross Abstract: Reinforcement learning (RL) has become essential for unlocking advanced reasoning capabilities in large language models (LLMs). RL workflows i

RoboAgent: Chaining Basic Capabilities for Embodied Task Planning

SafetyDGX agent

arXiv:2604.07774v1 Announce Type: cross Abstract: This paper focuses on embodied task planning, where an agent acquires visual observations from the environment and executes atomic actions to accompli

Robust Multi-Agent Target Tracking in Intermittent Communication Environments via Analytical Belief Merging

AgentsDGX agent

arXiv:2604.07575v1 Announce Type: new Abstract: Autonomous multi-agent target tracking in GPS-denied and communication-restricted environments (e.g., underwater exploration, subterranean search and re

Robust support vector model based on bounded asymmetric elastic net loss for binary classification

Model ReleasesDGX agent

arXiv:2603.06257v2 Announce Type: replace-cross Abstract: In this paper, we propose a novel bounded asymmetric elastic net (L_{baen}) loss function and combine it with the support vector machine (SV

Robustness Risk of Conversational Retrieval: Identifying and Mitigating Noise Sensitivity in Qwen3-Embedding Model

Model ReleasesDGX agent

arXiv:2604.06176v1 Announce Type: cross Abstract: We present an empirical study of embedding-based retrieval under realistic conversational settings, where queries are short, dialogue-like, and weakly

Rocket Report: Chinese version of Falcon 9 fails; Artemis depends on rapid heavy lift

IndustryDGX agent

A weekly industry rocket news roundup highlighted three key developments: a Chinese Falcon 9-style reusable rocket suffered a launch failure, underscoring ongoing challenges facing China's growing ...

RoSHI: A Versatile Robot-oriented Suit for Human Data In-the-Wild

SafetyDGX agent

arXiv:2604.07331v1 Announce Type: cross Abstract: Scaling up robot learning will likely require human data containing rich and long-horizon interactions in the wild. Existing approaches for collecting

Rotation Equivariant Convolutions in Deformable Registration of Brain MRI

SafetyDGX agent

arXiv:2604.08034v1 Announce Type: new Abstract: Image registration is a fundamental task that aligns anatomical structures between images. While CNNs perform well, they lack rotation equivariance - a

RPM-Net Reciprocal Point MLP Network for Unknown Network Security Threat Detection

TutorialsDGX agent

arXiv:2604.06638v1 Announce Type: cross Abstract: Effective detection of unknown network security threats in multi-class imbalanced environments is critical for maintaining cyberspace security. Curren

RQR3D: Reparametrizing the regression targets for BEV-based 3D object detection

AgentsDGX agent

arXiv:2505.17732v2 Announce Type: replace Abstract: Accurate, fast, and reliable 3D perception is essential for autonomous driving. Recently, bird's-eye view (BEV)-based perception approaches have eme

$S^3$: Stratified Scaling Search for Test-Time in Diffusion Language Models

ResearchDGX agent

arXiv:2604.06260v1 Announce Type: cross Abstract: Test-time scaling investigates whether a fixed diffusion language model (DLM) can generate better outputs when given more inference compute, without a

Safe Large-Scale Robust Nonlinear MPC in Milliseconds via Reachability-Constrained System Level Synthesis on the GPU

SafetyDGX agent

arXiv:2604.07644v1 Announce Type: new Abstract: We present GPU-SLS, a GPU-parallelized framework for safe, robust nonlinear model predictive control (MPC) that scales to high-dimensional uncertain rob

SALLIE: Safeguarding Against Latent Language & Image Exploits

Model ReleasesDGX agent

arXiv:2604.06247v1 Announce Type: cross Abstract: Large Language Models (LLMs) and Vision-Language Models (VLMs) remain highly vulnerable to textual and visual jailbreaks, as well as prompt injections

Sampling-Aware 3D Spatial Analysis in Multiplexed Imaging

ResearchDGX agent

arXiv:2604.07890v1 Announce Type: new Abstract: Highly multiplexed microscopy enables rich spatial characterization of tissues at single-cell resolution, yet most analyses rely on two-dimensional sect

Sandbox SDK adds file permission control

ToolsDGX agent

Vercel Sandbox SDK version 1.9.0 introduces native file permission control within the `writeFiles` API, allowing developers to set Unix-style permissions by passing a `mode` property in a single op...

SANDO: Safe Autonomous Trajectory Planning for Dynamic Unknown Environments

Model ReleasesDGX agent

arXiv:2604.07599v1 Announce Type: new Abstract: SANDO is a safe trajectory planner for 3D dynamic unknown environments, where obstacle locations and motions are unknown a priori and a collision-free p

SAT: Balancing Reasoning Accuracy and Efficiency with Stepwise Adaptive Thinking

ResearchDGX agent

arXiv:2604.07922v1 Announce Type: cross Abstract: Large Reasoning Models (LRMs) have revolutionized complex problem-solving, yet they exhibit a pervasive 'overthinking', generating unnecessarily long

SAT: Selective Aggregation Transformer for Image Super-Resolution

Local AiDGX agent

arXiv:2604.07994v1 Announce Type: new Abstract: Transformer-based approaches have revolutionized image super-resolution by modeling long-range dependencies. However, the quadratic computational comple

Say Something Else: Rethinking Contextual Privacy as Information Sufficiency

ResearchDGX agent

arXiv:2604.06409v1 Announce Type: cross Abstract: LLM agents increasingly draft messages on behalf of users, yet users routinely overshare sensitive information and disagree on what counts as private.

SBBTS: A Unified Schrodinger-Bass Framework for Synthetic Financial Time Series

ResearchDGX agent

arXiv:2604.07159v1 Announce Type: new Abstract: We study the problem of generating synthetic time series that reproduce both marginal distributions and temporal dynamics, a central challenge in financ

Scal3R: Scalable Test-Time Training for Large-Scale 3D Reconstruction

Local AiDGX agent

arXiv:2604.08542v1 Announce Type: new Abstract: This paper addresses the task of large-scale 3D scene reconstruction from long video sequences. Recent feed-forward reconstruction models have shown pro

Scaling-Aware Data Selection for End-to-End Autonomous Driving Systems

AgentsDGX agent

arXiv:2604.08366v1 Announce Type: cross Abstract: Large-scale deep learning models for physical AI applications depend on diverse training data collection efforts. These models and correspondingly, th

← Previous
1…13681369137013711372…1387
Next →