AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,020
  • Agents7,759
  • Applications5,540
  • Concepts5
  • Hardware1,925
  • Industry6,204
  • Local Ai5,102
  • Model Releases24,783
  • Research20,783
  • Safety13,742
  • Syntheses17
  • Tools1,680
  • Tutorials3,480

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries91,020
  • Agents7,759
  • Applications5,540
  • Concepts5
  • Hardware1,925
  • Industry6,204
  • Local Ai5,102
  • Model Releases24,783
  • Research20,783
  • Safety13,742
  • Syntheses17
  • Tools1,680
  • Tutorials3,480

Source
Human
91,020Total entries
1Added by human
91,019Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
91,019 results
10 Jun 2026

QDepth-VLA: Quantized Depth Prediction as Auxiliary Supervision for Vision-Language-Action Models

TutorialsDGX agent

arXiv:2510.14836v3 Announce Type: replace Abstract: Spatial perception and reasoning are crucial for Vision-Language-Action (VLA) models to accomplish fine-grained manipulation tasks. However, existin

QSplitFL: Capability Aware Deep Q-Learning for Optimal Split Point Selection in Split Federated Learning

ResearchDGX agent

arXiv:2606.09869v1 Announce Type: cross Abstract: Federated Learning (FL) combined with Split Learning (SL) is a privacy preserving paradigm that enables training deep neural networks (DNNs) on resour

Quality Is Not a Safety Proxy Under Quantization

Model ReleasesDGX agent

arXiv:2606.10154v1 Announce Type: new Abstract: Quantized checkpoints are often screened first with quality metrics and only later, if at all, with direct safety tests. This paper audits that shortcut

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Quantifying Perception-Based Student Success with Generative AI: An Exploratory Monte Carlo Simulation

ApplicationsDGX agent

arXiv:2507.01062v4 Announce Type: replace-cross Abstract: Generative artificial intelligence (GenAI) tools such as ChatGPT have attracted growing attention in higher education, particularly in relatio

Quantifying Uncertainty in AI Visibility: A Statistical Framework for Generative Search Measurement

Model ReleasesDGX agent

arXiv:2603.08924v2 Announce Type: replace-cross Abstract: AI-powered answer engines are inherently non-deterministic: identical queries submitted at different times can produce different responses and

Question: tech market is down about 10% over the last week. Will that affect SpaceX’s IPO? Or will everything proceed as planned, since it i…

SafetyDGX agent

A question posed on X regarding whether a recent 10% decline in the tech market will impact SpaceX's planned IPO timing and execution. The post, from AI researcher and entrepreneur Gary Marcus, raises

Quo Vadis, Visual In-Context Learning? A Unified Benchmark Across Domains and Tasks

Model ReleasesDGX agent

arXiv:2606.10967v1 Announce Type: new Abstract: Visual in-context learning has been proposed as a pathway towards dynamic models that can generate predictions based on a provided context and thereby c

Quoting Jeremy Howard

Model ReleasesDGX agent

Easy solution to slow down recursive AI self improvement: The lab with the top-ranked model must agree THEY must not use it for working on frontier AI But everyone else should have access to it. By de

Racist comments targeting politicians tripled since Meta relaxed its rules

IndustryDGX agent

After Meta relaxed its content moderation policies in January 2025, analysis of nearly 8 million Facebook comments showed that abusive and racist comments targeting lawmakers from both parties tripled

Range Penalization: Theoretical Insights with Applications in Federated Learning

ResearchDGX agent

arXiv:2606.10916v1 Announce Type: cross Abstract: This paper introduces range regularization for federated learning with linear systematic components to enhance statistical accuracy and induce cross-c

Rank Collapse, Fixed Points, and the Renormalization Group Structure of MLP Residual Networks

Model ReleasesDGX agent

arXiv:2606.10324v1 Announce Type: new Abstract: The analogy between deep neural network forward passes and renormalization group (RG) flows has been repeatedly noted in the literature, but existing tr

RankLLM: Weighted Ranking of LLMs by Quantifying Question Difficulty

ResearchDGX agent

arXiv:2602.12424v2 Announce Type: replace-cross Abstract: Benchmarks establish a standardized evaluation framework to systematically assess the performance of large language models (LLMs), facilitatin

RAPTOR: Rapid Aerial Pickup and Transport of Objects by Robots

ApplicationsDGX agent

arXiv:2203.03018v3 Announce Type: replace Abstract: Rapid aerial grasping through robots can lead to many applications that utilize fast and dynamic picking and placing of objects. Rigid grippers trad

RAT: Reference-Augmented Training for ASV Anti-Spoofing

Model ReleasesDGX agent

arXiv:2606.10908v1 Announce Type: cross Abstract: We introduce a spoofing countermeasure architecture conditioned on speaker-reference recordings, but observe that it converges to a solution that effe

READER: Robust Evidence-based Authorship Decoding via Extracted Representations

Model ReleasesDGX agent

arXiv:2606.10794v1 Announce Type: new Abstract: As agentic applications increasingly route user tasks through official and third-party LLM APIs, provenance becomes an operational question: which model

REAL: A Reasoning-Enhanced Graph Framework for Long-Term Memory Management of LLMs

Model ReleasesDGX agent

arXiv:2606.10694v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly expected to interact with users over long time horizons. However, due to their finite context window, LLMs

Really enjoyed reading the Microsoft MAI-Thinking-1 'Building a Hill Climbing Machine' paper. Amazing they publicly released all the info ne…

Model ReleasesDGX agent

Really enjoyed reading the Microsoft MAI-Thinking-1 'Building a Hill Climbing Machine' paper. Amazing they publicly released all the info needed to train a frontier model, down to hparams. I also thou

really good point. who else?

SafetyDGX agent

really good point. who else? BREAKING; Bill Gates just told Congress that Jeffrey Epstein had sought to use Gates’ affairs in an effort to blackmail him. So the question now is, if he did this to Bill

RealMath-Eval: Why SOTA Judges Struggle with Real Human Reasoning

Model ReleasesDGX agent

arXiv:2606.10254v1 Announce Type: new Abstract: While Large Language Models (LLMs) have achieved near-perfect performance in solving high-school mathematics, their ability to evaluate the diverse reas

ReasonAlloc: Hierarchical Decoding-Time KV Cache Budget Allocation for Reasoning Models

Model ReleasesDGX agent

arXiv:2606.11164v1 Announce Type: new Abstract: Long chain-of-thought (CoT) trajectories in large language model (LLM) reasoning cause severe inference bottlenecks due to rapid key-value (KV) cache gr

Reasoning or Memorization? Direction-Aware Diversity Exploration in LLM Reinforcement Learning

SafetyDGX agent

arXiv:2606.10346v1 Announce Type: new Abstract: Reinforcement learning has become a key paradigm for eliciting reasoning abilities in large language models, where exploration is crucial for discoverin

Reasoning over Semantic IDs Enhances Generative Recommendation

SafetyDGX agent

arXiv:2603.23183v2 Announce Type: replace-cross Abstract: Recent advances in generative recommendation have leveraged pretrained LLMs by formulating sequential recommendation as autoregressive generat

Recalling Too Well: Sycophancy Evaluation and Mitigation in Memory-Augmented Models

Model ReleasesDGX agent

arXiv:2606.10949v1 Announce Type: new Abstract: Persistent memory systems promise to make LLMs more helpful by storing user beliefs over time. We show they also make models less correct by systematica

Recoverable but Not Stationary:Local Linear Structures in Weights and Activations

Model ReleasesDGX agent

arXiv:2606.10929v1 Announce Type: cross Abstract: Task vectors, LoRA, activation steering, and random search around pretrained weights all suggest that learned behaviour can be controlled by linear di

Recovering the Zipfian Distribution in Unsupervised Term Discovery

SafetyDGX agent

arXiv:2606.10781v1 Announce Type: cross Abstract: Unsupervised term discovery involves segmenting unlabelled speech into word- or syllable-like units and clustering these into a lexicon of candidate t

recruiting two singer-researchers to stage a dramatic adaptiation of the Muon/Shampoo debate set to the tune of “Your Obedient Servant” from…

ToolsDGX agent

Swyx posted about recruiting singer-researchers to create a dramatic musical adaptation of the Muon/Shampoo debate, using the melody from 'Your Obedient Servant' (likely referencing the Hamilton music

RedAct: Redacting Agent Capability Traces for Procedural Skill Protection

Model ReleasesDGX agent

arXiv:2606.10813v1 Announce Type: cross Abstract: Users rely on execution traces to observe agent behavior, diagnose failures, and ensure accountability. These traces contain rich procedural detail, i

ReflectiChain: Epistemic Grounding in LLM-Driven World Models for Supply Chain Resilience

Model ReleasesDGX agent

arXiv:2606.10359v1 Announce Type: new Abstract: AI agents in supply chains face a fundamental epistemic gap: large language models (LLMs) interpret policies but lack physical grounding, while reinforc

Regimes: An Auditable, Held-Out-Gated Improvement Loop Demonstrated on LongMemEval with ActiveGraph

AgentsDGX agent

arXiv:2606.10241v1 Announce Type: new Abstract: Autonomous improvement loops are hard to trust because the improvement process is usually external scaffolding bolted onto the agent: failures go unlogg

Representation-Aware Advantage Estimation: Your Reward Model Provides More Than A Scalar Output

ResearchDGX agent

arXiv:2606.10528v1 Announce Type: cross Abstract: Current reinforcement learning from human feedback (RLHF) methods primarily rely on scalar rewards from a trained reward model (RM). While effective,

Representation Curriculum: Stagewise Training for Robust Ranking and Allocation

SafetyDGX agent

arXiv:2606.09891v1 Announce Type: cross Abstract: Ranking in digital marketplaces is a dynamic exposure-allocation mechanism: displayed items shape discovery trajectories and success events logged by

Resilient Navigation for Autonomous Farm Robots by Leveraging Jerk-Augmented Models with IMU-Only Disturbance Rejection

AgentsDGX agent

arXiv:2606.10971v1 Announce Type: new Abstract: Precise state estimation for navigation of autonomous agricultural robots is often compromised by sensor outages (GNSS/LiDAR/Visual) and high-frequency

Rethinking Embodied Navigation via Relational Inductive Bias

SafetyDGX agent

arXiv:2606.10348v1 Announce Type: new Abstract: Object navigation requires an agent to locate a target in an unknown environment through visual observations. Existing methods typically rely on open-vo

Rethinking the Flow-Based Gradual Domain Adaptation: A Semi-Dual Optimal Transport Perspective

ResearchDGX agent

arXiv:2602.01179v2 Announce Type: replace Abstract: Gradual domain adaptation (GDA) aims to mitigate domain shift by progressively adapting models from the source domain to the target domain via inter

Revisiting Metric Reliability for Fine-grained Evaluation of Machine Translation and Summarization in Indian Languages

Model ReleasesDGX agent

arXiv:2510.07061v2 Announce Type: replace Abstract: While automatic metrics drive progress in Machine Translation (MT) and Text Summarization (TS), existing metrics have been developed and validated a

Revisiting Positive Samples in Graph Contrastive Learning: From the Perspective of Message Passing

ResearchDGX agent

arXiv:2606.10284v1 Announce Type: new Abstract: Graph Contrastive Learning (GCL), which trains graph encoders by maximizing similarity between positive samples and minimizing it between negative ones,

Risk Comparisons in Linear Regression: Implicit Regularization Dominates Explicit Regularization

ResearchDGX agent

arXiv:2509.17251v2 Announce Type: replace-cross Abstract: Existing theory suggests that for linear regression problems categorized by capacity and source conditions, gradient descent (GD) is always mi

RKSC: Reasoning-Aware KV Cache Sharing and Confident Early Exit for Multi-Step LLM Inference

ResearchDGX agent

arXiv:2606.09937v1 Announce Type: cross Abstract: We introduce RKSC (Reasoning-Aware KV Cache Sharing), a training-free inference framework that eliminates two structural redundancies in multi-branch

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning

Model ReleasesDGX agent

arXiv:2510.14828v3 Announce Type: replace Abstract: Improving the reasoning capabilities of embodied agents is crucial for robots to complete complex human instructions in long-view manipulation tasks

RoboNaldo: Accurate, Stable and Powerful Humanoid Soccer Shooting via Motion-Guided Curriculum Reinforcement Learning

SafetyDGX agent

arXiv:2606.11092v1 Announce Type: cross Abstract: Elite humanoid soccer shooting requires whole-body stability, high-impulse whole-body interactions, and accuracy to targets. Motion tracking-driven re

RobotEQ: Transitioning from Passive Intelligence to Active Intelligence in Embodied AI

Model ReleasesDGX agent

arXiv:2605.06234v2 Announce Type: replace Abstract: Embodied AI is a prominent research topic in both academia and industry. Current research centers on completing tasks based on explicit user instruc

Robotic Nonprehensile Object Transportation with a Hanging Tray

ResearchDGX agent

arXiv:2606.10039v1 Announce Type: new Abstract: We consider the nonprehensile object transportation task known as the waiter's problem, in which a robot must move an object balanced on a tray from one

Robust Active Learning for Few-Shot Example Selection in Text-to-SQL

ResearchDGX agent

arXiv:2606.10125v1 Announce Type: cross Abstract: Few-shot example retrieval is the dominant paradigm for grounding large language models (LLMs) in domain-specific text-to-SQL systems. However, the qu

Robust Deep Reinforcement Learning Through Adversarial Attacks and Training : A Survey

AgentsDGX agent

arXiv:2403.00420v3 Announce Type: replace-cross Abstract: Deep Reinforcement Learning (DRL) is a subfield of machine learning for training autonomous agents that take sequential actions across complex

Robust Regression of General ReLUs with Queries

SafetyDGX agent

arXiv:2606.11130v1 Announce Type: new Abstract: We study the task of agnostically learning general (as opposed to homogeneous) ReLUs under the Gaussian distribution with respect to the squared loss. I

“Rockets are cool. There's no getting around that.” — Elon Musk

IndustryDGX agent

Elon Musk expresses enthusiasm for rocket technology in a casual social media post, reflecting his general passion for spaceflight and space exploration. The statement represents Musk's characteristic

Rod models in continuum and soft robot control: a review

ApplicationsDGX agent

arXiv:2407.05886v3 Announce Type: replace Abstract: Continuum and soft robots can transform automation tasks requiring compliant interaction in constrained or unstructured environments, including heal

Role-Agent: Bootstrapping LLM Agents via Dual-Role Evolution

SafetyDGX agent

arXiv:2606.10917v1 Announce Type: new Abstract: Although Large Language Model (LLM) agents have demonstrated strong performance on complex tasks, their learning is often limited by inefficient interac

ros2probe: Non-intrusive, Kernel-selective Observability for Robot Operating System 2 Middleware

ResearchDGX agent

arXiv:2606.10746v1 Announce Type: new Abstract: Robot Operating System 2 (ROS 2), the de facto standard middleware framework for robots, runs each robot as a graph of nodes communicating over the Data

Rotate2Think: Geometric Priming via Orthogonal Rotation to Improve Language Model Reasoning

Model ReleasesDGX agent

arXiv:2606.09873v1 Announce Type: cross Abstract: Reasoning models achieve strong performance on challenging tasks by generating explicit intermediate reasoning traces before producing a final answer.

Routing-Aware Expert Calibration for Machine Unlearning in Mixture-of-Experts Language Models

ResearchDGX agent

arXiv:2606.10338v1 Announce Type: cross Abstract: Machine unlearning is increasingly important for large language models, yet unlearning in Mixture-of-Experts (MoE) architectures remains underexplored

Run DiffusionGemma on NVIDIA for Developer-Ready, High-Throughput Text Generation

HardwareDGX agent

DiffusionGemma is an experimental open model built for exceptionally fast text generation that NVIDIA has optimized to run on GeForce RTX GPUs, RTX PRO, and DGX Spark systems. Rather than generating t

RunPod AI Hub - Public Beta 1.34 live

Local AiDGX agent

RunPod Hub is a centralized catalog of preconfigured AI repositories that you can browse, deploy, and share, optimized for RunPod's Serverless infrastructure to deploy in minutes. The platform include

SAFE: An LLM-as-Verifier Framework for Evidence-Grounded Multi-Hop Reasoning

Model ReleasesDGX agent

arXiv:2604.01993v2 Announce Type: replace-cross Abstract: Multi-hop QA benchmarks often reward Large Language Models (LLMs) for spurious correctness, where models reach correct answers through invalid

Sample Where You Struggle: Sharpening Base Model Reasoning via Entropy-Guided Power Sampling

Model ReleasesDGX agent

arXiv:2606.09926v1 Announce Type: cross Abstract: Sampling from the sequence-level power distribution p^alpha elicits RL-level reasoning from base language models without any parameter updates, but th

Sampling Triangulations and Calabi-Yau Threefolds with Autoregressive GNNs

HardwareDGX agent

arXiv:2605.27770v2 Announce Type: cross Abstract: We introduce `dualGNN', an autoregressive message-passing GNN for sampling fine, regular triangulations (FRTs) of convex polytopes. dualGNN operates o

SARM2: Multi-Task Stage Aware Reward Modeling for Self Improving Robotic Manipulation

Model ReleasesDGX agent

arXiv:2606.10305v1 Announce Type: new Abstract: Fine-tuning vision-language-action (VLA) policies for long-horizon manipulation still relies heavily on behavior cloning, which requires costly high-qua

Say what you will about OpenAI but at least their employees debate their positions in public

TutorialsDGX agent

This post likely praises OpenAI's organizational culture for encouraging public discourse and debate among employees about company decisions and positions. The statement suggests that OpenAI demonstra

SCAIL-2: Unifying Controlled Character Animation with End-to-end In-Context Conditioning

Model ReleasesDGX agent

arXiv:2606.10804v1 Announce Type: new Abstract: Controlled character animation requires transferring motion from a driving sequence to a reference character. Prior works heavily rely on intermediate r

Scaling AI Through Data Fluency

IndustryDGX agent

This Databricks blog post discusses how organizations can scale their AI initiatives by developing data fluency—the ability to effectively understand, manage, and leverage data across the enterprise.

← Previous
1…626627628629630…1517
Next →