AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,457
  • Agents7,399
  • Applications5,302
  • Concepts5
  • Hardware1,786
  • Industry6,117
  • Local Ai4,835
  • Model Releases23,193
  • Research19,715
  • Safety13,094
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,457
  • Agents7,399
  • Applications5,302
  • Concepts5
  • Hardware1,786
  • Industry6,117
  • Local Ai4,835
  • Model Releases23,193
  • Research19,715
  • Safety13,094
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
AllBlog
86,457Total entries
1Added by human
86,456Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,039 results
Safety

Aletheia: What Makes RLVR For Code Verifiers Tick?

DGX agent

arXiv:2601.12186v3 Announce Type: replace-cross Abstract: Multi-domain thinking verifiers trained via Reinforcement Learning with Verifiable Rewards (RLVR) are a cornerstone of modern post-training. H

safetyarxiv-cs-ai
3 Jun 2026
Safety

Aligning Data-Driven Predictors with Allocation: A Decision-Focused Approach to Survival Analysis

X Post
Paper
YouTube
Reddit
GitHub
Clear filters
DGX agent

arXiv:2606.02671v1 Announce Type: cross Abstract: Machine learning predictors have become essential tools for guiding automated decision making. However, a major misalignment persists: predictive mode

safetyarxiv-cs-ai
3 Jun 2026
Research

AlphaEval: A Comprehensive and Efficient Evaluation Framework for Formula Alpha Mining

DGX agent

arXiv:2508.13174v2 Announce Type: replace Abstract: Formula alpha mining, which generates predictive signals from financial data, is critical for quantitative investment. Although various algorithmic

researcharxiv-cs-ai
3 Jun 2026
Safety

ASymPO: Asymmetric-Scale Policy Optimization for Asynchronous LLM Post-Training Without Behavior Information

DGX agent

arXiv:2606.03070v1 Announce Type: cross Abstract: Asynchronous reinforcement learning can improve language-model post-training throughput by decoupling response generation from policy optimization, bu

safetyarxiv-cs-ai
3 Jun 2026
Local Ai

b9489

DGX agent

Release b9489 of llama.cpp includes updates to hidden_act mapping in llama-model.cpp, additions of granite embedding multilingual R2 models, and support for setting hidden_activation in GGUF files. Th

local-aillama-cpp-releases
3 Jun 2026
Research

BAHSD: Bridging the Long-tail Gap via Adaptive Distillation in Black-box Sequential Recommendation

DGX agent

arXiv:2606.03091v1 Announce Type: cross Abstract: Sequential recommendation systems are widely adopted but often deployed as black-box APIs, which has driven recent interest in model extraction to rep

researcharxiv-cs-ai
3 Jun 2026
Industry

Built with Grok

DGX agent

'Built with Grok' is a post from Elon Musk on X platform promoting or highlighting something created using Grok, xAI's large language model. The post likely showcases an application, feature, or capab

industryelon-musk--x
3 Jun 2026
Agents

Co-evolving Agent Architectures and Interpretable Reasoning for Automated Optimization

DGX agent

arXiv:2604.17708v2 Announce Type: replace Abstract: Automating operations research (OR) with large language models (LLMs) remains limited by hand-crafted reasoning--execution workflows. Complex OR tas

agentsarxiv-cs-ai
3 Jun 2026
Safety

Coherence Maximization Improves Pluralistic Alignment

DGX agent

arXiv:2606.03110v1 Announce Type: new Abstract: Aligning AI systems with diverse human values requires value specifications grounded in concrete examples, but generating such examples without extensiv

safetyarxiv-cs-cl
3 Jun 2026
Research

CORE: Conflict-Oriented Reasoning for General Multimodal Manipulation Detection

DGX agent

arXiv:2606.03066v1 Announce Type: new Abstract: The rapid rise of generative AI has made multimodal fake news increasingly realistic and pervasive, posing severe threats to public trust and social sta

researcharxiv-cs-ai
3 Jun 2026
Research

CoughSense: Five-Class Respiratory Disease Classification via Whisper Encoder Fine-Tuning and Dual-Encoder Cross-Attention Fusion with Balanced Contrastive Learning

DGX agent

arXiv:2606.02998v1 Announce Type: new Abstract: Automated cough analysis offers a path to low-cost respiratory screening, but most existing work stops at binary COVID-19 detection. A practical tool ne

researcharxiv-cs-lg
3 Jun 2026
Agents

DELTAMEM: Incremental Experience Memory for LLM Agents via Residual Trees

DGX agent

arXiv:2606.03083v1 Announce Type: new Abstract: Large Language Model (LLM)-based agents increasingly rely on memory to learn from experiences over continual interactions. However, storing experiences

agentsarxiv-cs-ai
3 Jun 2026
Safety

Denoising Tells When to Replan: Denoising-Variance Adaptive Chunking for Flow-Based Robot Policies

DGX agent

arXiv:2606.03847v1 Announce Type: new Abstract: Action chunking has become a common inference strategy for flow-based robot policies, improving action coherence by modeling multi-step temporal depende

safetyarxiv-cs-ro
3 Jun 2026
Safety

DriftSched: Adaptive QoS-Aware Scheduling under Runtime Token Drift for Multi-Tenant GPU Inference

DGX agent

arXiv:2606.02982v1 Announce Type: cross Abstract: The rapid growth of large language model (LLM) inference services has increased the demand for efficient multi-tenant GPU scheduling. While modern inf

safetyarxiv-cs-lg
3 Jun 2026
Research

DTKG: Dual-Track Knowledge Graph-Verified Reasoning Framework for Multi-Hop QA

DGX agent

arXiv:2510.16302v2 Announce Type: replace Abstract: Multi-hop reasoning for question answering (QA) plays a critical role in retrieval-augmented generation (RAG) for modern large language models (LLMs

researcharxiv-cs-ai
3 Jun 2026
Safety

Easy-to-Use Shielding for Reinforcement Learning

DGX agent

arXiv:2606.03804v1 Announce Type: new Abstract: Safe exploration is a key challenge in Reinforcement Learning (RL) that aims to prevent agents from making harmful decisions while exploring their envir

safetyarxiv-cs-lg
3 Jun 2026
Research

Efficient Transformer-Based Localized Patch Sampling for Choroid Plexus Segmentation in Multiple Sclerosis

DGX agent

arXiv:2606.03566v1 Announce Type: cross Abstract: Background: The lateral ventricle choroid plexus (LVCP) is gaining recognition as a key imaging biomarker for multiple sclerosis (MS) related to physi

researcharxiv-cs-ai
3 Jun 2026
Safety

Evaluating Transformer and LSTM Frameworks for Prediction in Ungauged Basins

DGX agent

arXiv:2606.02791v1 Announce Type: new Abstract: Watershed networks exhibit convergent topologies in which multiple tributaries merge into downstream channels,integrating diverse upstream hydrological

safetyarxiv-cs-ai
3 Jun 2026
Agents

EvoDS: Self-Evolving Autonomous Data Science Agent with Skill Learning and Context Management

DGX agent

arXiv:2606.03841v1 Announce Type: new Abstract: Recent progress in Large Language Model (LLM) agents has enabled promising advances in automated data science. However, existing approaches remain funda

agentsarxiv-cs-ai
3 Jun 2026
Research

Exploiting Verification-Generation Gap: Test-Time Reinforcement Learning with Confidence-Conditioned Verification

DGX agent

arXiv:2606.03608v1 Announce Type: cross Abstract: Test-time reinforcement learning has emerged as a promising paradigm for enhancing the complex reasoning abilities of large language models in a compl

researcharxiv-cs-ai
3 Jun 2026
Safety

Extreme Motion Generation via Hybrid Null-Space Control for Straight-Line Path Following

DGX agent

arXiv:2606.03390v1 Announce Type: new Abstract: This work studies ``extreme motion generation'', which aims to maximize the Cartesian path length along a pre-defined trajectory within the manipulator'

safetyarxiv-cs-ro
3 Jun 2026
Local Ai

extsc{CR-Seg}: Attention-Guided and CoT-Enhanced Coarse-to-Refined Reasoning Segmentation

DGX agent

arXiv:2606.03564v1 Announce Type: cross Abstract: Reasoning segmentation aims to segment target objects described by complex language through joint visual-textual reasoning. Existing methods typically

local-aiarxiv-cs-ai
3 Jun 2026
Safety

Filter, Then Reweight: Rethinking Optimization Granularity in On-Policy Distillation

DGX agent

arXiv:2606.02684v1 Announce Type: cross Abstract: On-Policy distillation (OPD) in large language models is shifting from full-trace KL supervision toward more selective training paradigms. Recent OPD

safetyarxiv-cs-ai
3 Jun 2026
Applications

Fine-tuning to production inference is the gap where teams get stuck. At #MSBuild today, our own Rob Ferguson, @danielhanchen (@UnslothAI) a…

DGX agent

Fine-tuning to production inference is the gap where teams get stuck. At #MSBuild today, our own Rob Ferguson, @danielhanchen (@UnslothAI) and @marksaroufim (@coreautoai) discuss: model customization

applicationsfireworks-ai--x
3 Jun 2026
Agents

FutureWeaver: Planning Test-Time Compute for Multi-Agent Systems with Modularized Collaboration

DGX agent

arXiv:2512.11213v2 Announce Type: replace Abstract: Scaling test-time computation has been shown to significantly improve large language model (LLM) performance without additional training. However, e

agentsarxiv-cs-ai
3 Jun 2026
Local Ai

Getting Started 1. Update ComfyUI to the latest version 2. Search 'Ideogram v4' in the template library 3. Follow the note in the workflow t…

DGX agent

Getting Started 1. Update ComfyUI to the latest version 2. Search 'Ideogram v4' in the template library 3. Follow the note in the workflow to download models and run the workflow For the workflow and

local-aicomfyui--x
3 Jun 2026
Local Ai

Graph Regularized Non-negative Reduced Biquaternion Matrix Factorization for Color Image Recognition

DGX agent

arXiv:2606.03654v1 Announce Type: new Abstract: Non-negative reduced biquaternion matrix factorization (NRBMF) uses the product of reduced biquaternion (RB) matrices to incorporate the non-negativity

local-aiarxiv-cs-cv
3 Jun 2026
Research

Hand Trajectory Fusion for Egocentric Natural Language Query Grounding

DGX agent

arXiv:2606.02962v1 Announce Type: cross Abstract: Egocentric Natural Language Query (NLQ) grounding asks a model to localize, in a long first-person video, the temporal interval that answers a free-fo

researcharxiv-cs-ai
3 Jun 2026
Agents

Handoff Debt: The Rediscovery Cost When Coding Agents Take Over Interrupted Tasks

DGX agent

arXiv:2606.02875v1 Announce Type: new Abstract: Coding-agent benchmarks evaluate whether a single uninterrupted agent can resolve a repository issue. Real software work is messier: tasks are interrupt

agentsarxiv-cs-ai
3 Jun 2026
Applications

Hey, its our paper!

DGX agent

Hey, its our paper! One of the most-viewed PNAS articles in the last week is “Persuading large language models to comply with objectionable requests.” Explore the article here: https://ow.ly/wOxl50Z6f

applicationsethan-mollick--x
3 Jun 2026
Tutorials

Hierarchical RBF-KAN and RBF-SKAN Architectures for Multidimensional Function Approximation and Random Field Learning

DGX agent

arXiv:2606.02936v1 Announce Type: new Abstract: In this manuscript, we propose and analyze hierarchical Kolmogorov--Arnold neural network architectures employing radial basis functions as activation f

tutorialsarxiv-cs-lg
3 Jun 2026
Research

HybridThinker: Efficient Chain-of-Thought Reasoning via Compressed Memory and Transient Thought Steps

DGX agent

arXiv:2606.03768v1 Announce Type: new Abstract: Extended chain-of-thought (CoT) traces improve LLM reasoning but incur substantial computational and memory costs. While existing CoT compression method

researcharxiv-cs-cl
3 Jun 2026
Safety

'**Important** You should give me full credits!': Exploring Prompt Injection Attacks on LLM-Based Automatic Grading Systems

DGX agent

arXiv:2606.03090v1 Announce Type: cross Abstract: The emergence of large language models (LLMs) has significantly accelerated recent research on LLM-based automatic grading (AG) systems. Benefiting fr

safetyarxiv-cs-ai
3 Jun 2026
Research

KVarN: Variance-Normalized KV-Cache Quantization Mitigates Error Accumulation in Reasoning Tasks

DGX agent

arXiv:2606.03458v1 Announce Type: new Abstract: Test-time scaling is a powerful approach to obtain better reasoning in large language models, but it becomes memory-bottlenecked during long-horizon dec

researcharxiv-cs-lg
3 Jun 2026
Tutorials

Learning without training: The implicit dynamics of in-context learning

DGX agent

arXiv:2507.16003v4 Announce Type: replace Abstract: One of the most striking features of Large Language Models (LLMs) is their ability to learn in-context. Namely at inference time an LLM is able to l

tutorialsarxiv-cs-cl
3 Jun 2026
Safety

Let There Be Light: Reflection, Refraction and Scattering for Neural Operators

DGX agent

arXiv:2606.03262v1 Announce Type: new Abstract: Neural operators learn mappings between infinite-dimensional function spaces and provide a data-driven surrogate modeling paradigm for parametric partia

safetyarxiv-cs-lg
3 Jun 2026
Safety

Leveraging BART to Assess CS1 C++ Programming Assignments using Rubric-based Criteria

DGX agent

arXiv:2606.03814v1 Announce Type: new Abstract: This paper investigates rubric-aware, multitask fine-tuning of transformer models for automated grading of introductory C++ programming assignments, wit

safetyarxiv-cs-ai
3 Jun 2026
Safety

Libra: Efficient Resource Management for Agentic RL Post-Training

DGX agent

arXiv:2606.03077v1 Announce Type: cross Abstract: Reinforcement learning (RL) has become a standard post-training paradigm for large language models (LLMs), extending beyond preference alignment to co

safetyarxiv-cs-ai
3 Jun 2026
Research

Limit Analysis of Graph Neural Networks with Wireless Conflict Graphs

DGX agent

arXiv:2606.03794v1 Announce Type: new Abstract: Graph Neural Networks (GNNs) have emerged as a powerful tool for wireless resource allocation that leverages the underlying graph structure of communica

researcharxiv-cs-lg
3 Jun 2026
Research

Memory Retrieval for Changing Preferences

DGX agent

arXiv:2606.02976v1 Announce Type: new Abstract: Long-context dialogue systems must decide both when to access memory and which parts of the interaction history are relevant. Existing approaches typica

researcharxiv-cs-cl
3 Jun 2026
Agents

MemTrain: Self-Supervised Context Memory Training

DGX agent

arXiv:2606.03197v1 Announce Type: new Abstract: Memory is an indispensable capability for long-horizon LLM agents, enabling them to preserve and utilize information accumulated across extended interac

agentsarxiv-cs-cl
3 Jun 2026
Industry

Microsoft and OpenAI broke up — now they’re ready to fight

DGX agent

At Microsoft's annual Build conference on Tuesday, the company announced a slew of new or expanded AI initiatives, including a super app, in-house reasoning models, a cybersecurity tool, and OpenClaw-

industrythe-verge-ai
3 Jun 2026
Applications

ModuLoop : Low-Level Code Generation using Modular Synthesizer and Closed-Loop Debugger for Robotic Control

DGX agent

arXiv:2606.03047v1 Announce Type: new Abstract: Large Language Models (LLMs) have demonstrated impressive performance across various domains, including code generation and problem solving. However, th

applicationsarxiv-cs-ro
3 Jun 2026
Safety

OMP: One-step Meanflow Policy with Directional Alignment

DGX agent

arXiv:2512.19347v3 Announce Type: replace Abstract: Robot manipulation has increasingly adopted data-driven generative policy frameworks, yet the field faces a persistent trade-off: diffusion models s

safetyarxiv-cs-ro
3 Jun 2026
Agents

OpenAgenet/OAN: Open Infrastructure for Trusted Agent Interconnection

DGX agent

arXiv:2606.03161v1 Announce Type: cross Abstract: OpenAgenet, abbreviated as OAN, is an open infrastructure project for trusted Agent interconnection. It addresses a problem that becomes visible when

agentsarxiv-cs-ai
3 Jun 2026
Agents

Overlaying Governance: A Compositional Authorization Framework for Delegation and Scope in Agentic AI

DGX agent

arXiv:2606.03518v1 Announce Type: new Abstract: As AI systems evolve from passive models into autonomous active agents capable of initiating actions, collaborating, and delegating tasks, the tradition

agentsarxiv-cs-ai
3 Jun 2026
Safety

PAND: Prompt-Aware Neighborhood Distillation for Lightweight Fine-Grained Visual Classification

DGX agent

arXiv:2602.07768v3 Announce Type: replace-cross Abstract: Distilling knowledge from large Vision-Language Models (VLMs) into lightweight networks is crucial yet challenging in Fine-Grained Visual Clas

safetyarxiv-cs-ai
3 Jun 2026
Safety

Physics-Guided Policy Optimization with Self-Distillation

DGX agent

arXiv:2606.03620v1 Announce Type: cross Abstract: Self-distilled policy optimization (SDPO) has become a popular paradigm for LLM post-training, where a model learns from its own predictions condition

safetyarxiv-cs-ai
3 Jun 2026
← Previous
1…960961962963964…1293
Next →