AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,428
  • Agents7,398
  • Applications5,301
  • Concepts5
  • Hardware1,785
  • Industry6,113
  • Local Ai4,833
  • Model Releases23,177
  • Research19,713
  • Safety13,092
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,428
  • Agents7,398
  • Applications5,301
  • Concepts5
  • Hardware1,785
  • Industry6,113
  • Local Ai4,833
  • Model Releases23,177
  • Research19,713
  • Safety13,092
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
86,428Total entries
1Added by human
86,427Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,015 results
Research

Exploiting Verification-Generation Gap: Test-Time Reinforcement Learning with Confidence-Conditioned Verification

DGX agent

arXiv:2606.03608v1 Announce Type: cross Abstract: Test-time reinforcement learning has emerged as a promising paradigm for enhancing the complex reasoning abilities of large language models in a compl

researcharxiv-cs-ai
3 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Extreme Motion Generation via Hybrid Null-Space Control for Straight-Line Path Following

DGX agent

arXiv:2606.03390v1 Announce Type: new Abstract: This work studies ``extreme motion generation'', which aims to maximize the Cartesian path length along a pre-defined trajectory within the manipulator'

safetyarxiv-cs-ro
3 Jun 2026
Local Ai

extsc{CR-Seg}: Attention-Guided and CoT-Enhanced Coarse-to-Refined Reasoning Segmentation

DGX agent

arXiv:2606.03564v1 Announce Type: cross Abstract: Reasoning segmentation aims to segment target objects described by complex language through joint visual-textual reasoning. Existing methods typically

local-aiarxiv-cs-ai
3 Jun 2026
Safety

Filter, Then Reweight: Rethinking Optimization Granularity in On-Policy Distillation

DGX agent

arXiv:2606.02684v1 Announce Type: cross Abstract: On-Policy distillation (OPD) in large language models is shifting from full-trace KL supervision toward more selective training paradigms. Recent OPD

safetyarxiv-cs-ai
3 Jun 2026
Applications

Fine-tuning to production inference is the gap where teams get stuck. At #MSBuild today, our own Rob Ferguson, @danielhanchen (@UnslothAI) a…

DGX agent

Fine-tuning to production inference is the gap where teams get stuck. At #MSBuild today, our own Rob Ferguson, @danielhanchen (@UnslothAI) and @marksaroufim (@coreautoai) discuss: model customization

applicationsfireworks-ai--x
3 Jun 2026
Agents

FutureWeaver: Planning Test-Time Compute for Multi-Agent Systems with Modularized Collaboration

DGX agent

arXiv:2512.11213v2 Announce Type: replace Abstract: Scaling test-time computation has been shown to significantly improve large language model (LLM) performance without additional training. However, e

agentsarxiv-cs-ai
3 Jun 2026
Local Ai

Getting Started 1. Update ComfyUI to the latest version 2. Search 'Ideogram v4' in the template library 3. Follow the note in the workflow t…

DGX agent

Getting Started 1. Update ComfyUI to the latest version 2. Search 'Ideogram v4' in the template library 3. Follow the note in the workflow to download models and run the workflow For the workflow and

local-aicomfyui--x
3 Jun 2026
Local Ai

Graph Regularized Non-negative Reduced Biquaternion Matrix Factorization for Color Image Recognition

DGX agent

arXiv:2606.03654v1 Announce Type: new Abstract: Non-negative reduced biquaternion matrix factorization (NRBMF) uses the product of reduced biquaternion (RB) matrices to incorporate the non-negativity

local-aiarxiv-cs-cv
3 Jun 2026
Research

Hand Trajectory Fusion for Egocentric Natural Language Query Grounding

DGX agent

arXiv:2606.02962v1 Announce Type: cross Abstract: Egocentric Natural Language Query (NLQ) grounding asks a model to localize, in a long first-person video, the temporal interval that answers a free-fo

researcharxiv-cs-ai
3 Jun 2026
Agents

Handoff Debt: The Rediscovery Cost When Coding Agents Take Over Interrupted Tasks

DGX agent

arXiv:2606.02875v1 Announce Type: new Abstract: Coding-agent benchmarks evaluate whether a single uninterrupted agent can resolve a repository issue. Real software work is messier: tasks are interrupt

agentsarxiv-cs-ai
3 Jun 2026
Applications

Hey, its our paper!

DGX agent

Hey, its our paper! One of the most-viewed PNAS articles in the last week is “Persuading large language models to comply with objectionable requests.” Explore the article here: https://ow.ly/wOxl50Z6f

applicationsethan-mollick--x
3 Jun 2026
Tutorials

Hierarchical RBF-KAN and RBF-SKAN Architectures for Multidimensional Function Approximation and Random Field Learning

DGX agent

arXiv:2606.02936v1 Announce Type: new Abstract: In this manuscript, we propose and analyze hierarchical Kolmogorov--Arnold neural network architectures employing radial basis functions as activation f

tutorialsarxiv-cs-lg
3 Jun 2026
Research

HybridThinker: Efficient Chain-of-Thought Reasoning via Compressed Memory and Transient Thought Steps

DGX agent

arXiv:2606.03768v1 Announce Type: new Abstract: Extended chain-of-thought (CoT) traces improve LLM reasoning but incur substantial computational and memory costs. While existing CoT compression method

researcharxiv-cs-cl
3 Jun 2026
Safety

'**Important** You should give me full credits!': Exploring Prompt Injection Attacks on LLM-Based Automatic Grading Systems

DGX agent

arXiv:2606.03090v1 Announce Type: cross Abstract: The emergence of large language models (LLMs) has significantly accelerated recent research on LLM-based automatic grading (AG) systems. Benefiting fr

safetyarxiv-cs-ai
3 Jun 2026
Research

KVarN: Variance-Normalized KV-Cache Quantization Mitigates Error Accumulation in Reasoning Tasks

DGX agent

arXiv:2606.03458v1 Announce Type: new Abstract: Test-time scaling is a powerful approach to obtain better reasoning in large language models, but it becomes memory-bottlenecked during long-horizon dec

researcharxiv-cs-lg
3 Jun 2026
Tutorials

Learning without training: The implicit dynamics of in-context learning

DGX agent

arXiv:2507.16003v4 Announce Type: replace Abstract: One of the most striking features of Large Language Models (LLMs) is their ability to learn in-context. Namely at inference time an LLM is able to l

tutorialsarxiv-cs-cl
3 Jun 2026
Safety

Let There Be Light: Reflection, Refraction and Scattering for Neural Operators

DGX agent

arXiv:2606.03262v1 Announce Type: new Abstract: Neural operators learn mappings between infinite-dimensional function spaces and provide a data-driven surrogate modeling paradigm for parametric partia

safetyarxiv-cs-lg
3 Jun 2026
Safety

Leveraging BART to Assess CS1 C++ Programming Assignments using Rubric-based Criteria

DGX agent

arXiv:2606.03814v1 Announce Type: new Abstract: This paper investigates rubric-aware, multitask fine-tuning of transformer models for automated grading of introductory C++ programming assignments, wit

safetyarxiv-cs-ai
3 Jun 2026
Safety

Libra: Efficient Resource Management for Agentic RL Post-Training

DGX agent

arXiv:2606.03077v1 Announce Type: cross Abstract: Reinforcement learning (RL) has become a standard post-training paradigm for large language models (LLMs), extending beyond preference alignment to co

safetyarxiv-cs-ai
3 Jun 2026
Research

Limit Analysis of Graph Neural Networks with Wireless Conflict Graphs

DGX agent

arXiv:2606.03794v1 Announce Type: new Abstract: Graph Neural Networks (GNNs) have emerged as a powerful tool for wireless resource allocation that leverages the underlying graph structure of communica

researcharxiv-cs-lg
3 Jun 2026
Research

Memory Retrieval for Changing Preferences

DGX agent

arXiv:2606.02976v1 Announce Type: new Abstract: Long-context dialogue systems must decide both when to access memory and which parts of the interaction history are relevant. Existing approaches typica

researcharxiv-cs-cl
3 Jun 2026
Agents

MemTrain: Self-Supervised Context Memory Training

DGX agent

arXiv:2606.03197v1 Announce Type: new Abstract: Memory is an indispensable capability for long-horizon LLM agents, enabling them to preserve and utilize information accumulated across extended interac

agentsarxiv-cs-cl
3 Jun 2026
Industry

Microsoft and OpenAI broke up — now they’re ready to fight

DGX agent

At Microsoft's annual Build conference on Tuesday, the company announced a slew of new or expanded AI initiatives, including a super app, in-house reasoning models, a cybersecurity tool, and OpenClaw-

industrythe-verge-ai
3 Jun 2026
Applications

ModuLoop : Low-Level Code Generation using Modular Synthesizer and Closed-Loop Debugger for Robotic Control

DGX agent

arXiv:2606.03047v1 Announce Type: new Abstract: Large Language Models (LLMs) have demonstrated impressive performance across various domains, including code generation and problem solving. However, th

applicationsarxiv-cs-ro
3 Jun 2026
Safety

OMP: One-step Meanflow Policy with Directional Alignment

DGX agent

arXiv:2512.19347v3 Announce Type: replace Abstract: Robot manipulation has increasingly adopted data-driven generative policy frameworks, yet the field faces a persistent trade-off: diffusion models s

safetyarxiv-cs-ro
3 Jun 2026
Agents

OpenAgenet/OAN: Open Infrastructure for Trusted Agent Interconnection

DGX agent

arXiv:2606.03161v1 Announce Type: cross Abstract: OpenAgenet, abbreviated as OAN, is an open infrastructure project for trusted Agent interconnection. It addresses a problem that becomes visible when

agentsarxiv-cs-ai
3 Jun 2026
Agents

Overlaying Governance: A Compositional Authorization Framework for Delegation and Scope in Agentic AI

DGX agent

arXiv:2606.03518v1 Announce Type: new Abstract: As AI systems evolve from passive models into autonomous active agents capable of initiating actions, collaborating, and delegating tasks, the tradition

agentsarxiv-cs-ai
3 Jun 2026
Safety

PAND: Prompt-Aware Neighborhood Distillation for Lightweight Fine-Grained Visual Classification

DGX agent

arXiv:2602.07768v3 Announce Type: replace-cross Abstract: Distilling knowledge from large Vision-Language Models (VLMs) into lightweight networks is crucial yet challenging in Fine-Grained Visual Clas

safetyarxiv-cs-ai
3 Jun 2026
Safety

Physics-Guided Policy Optimization with Self-Distillation

DGX agent

arXiv:2606.03620v1 Announce Type: cross Abstract: Self-distilled policy optimization (SDPO) has become a popular paradigm for LLM post-training, where a model learns from its own predictions condition

safetyarxiv-cs-ai
3 Jun 2026
Tools

probably the best reward function for reasoning efficiency i've seen

DGX agent

This post likely discusses an innovative reward function design that optimizes for reasoning efficiency in AI systems, possibly in the context of language models or reinforcement learning. The entry a

toolsswyx--x
3 Jun 2026
Research

R2DN: Scalable Parameterization of Contracting and Lipschitz Recurrent Deep Networks

DGX agent

arXiv:2504.01250v2 Announce Type: replace Abstract: This paper presents the Robust Recurrent Deep Network (R2DN), a scalable parameterization of robust recurrent neural networks for machine learning a

researcharxiv-cs-lg
3 Jun 2026
Local Ai

ResCLIP: Residual Attention for Training-free Dense Vision-language Inference

DGX agent

arXiv:2411.15851v2 Announce Type: replace Abstract: While vision-language models like CLIP have shown remarkable success in open-vocabulary tasks, their application is currently confined to image-leve

local-aiarxiv-cs-cv
3 Jun 2026
Research

Rethinking the Role of Tensor Decompositions in Post-Training LLM Compression

DGX agent

arXiv:2606.03465v1 Announce Type: cross Abstract: Post-training compression is essential for deploying large language models (LLMs) under tight resource constraints. Tensor decompositions have emerged

researcharxiv-cs-ai
3 Jun 2026
Agents

Say goodbye to month-end surprise invoices. LangSmith LLM Gateway lets you see your spend. Roll up your costs in real time by workspace, use…

DGX agent

LangSmith's LLM Gateway provides real-time cost visibility and spend tracking capabilities, allowing users to monitor their language model usage costs aggregated by workspace. This feature helps elimi

agentsharrison-chase--x
3 Jun 2026
Safety

See Less, Specify More: Visual Evidence Budgets for Generalizable VLAs

DGX agent

arXiv:2606.02735v1 Announce Type: cross Abstract: Generalization remains a central bottleneck for vision-language-action (VLA) models: under distractors, appearance shifts, and semantically similar ta

safetyarxiv-cs-ai
3 Jun 2026
Local Ai

SJD-PAC: Accelerating Speculative Jacobi Decoding via Proactive Drafting and Adaptive Continuation

DGX agent

arXiv:2603.18599v2 Announce Type: replace Abstract: Speculative Jacobi Decoding (SJD) offers a draft-model-free approach to accelerate autoregressive text-to-image synthesis. However, the high-entropy

local-aiarxiv-cs-cv
3 Jun 2026
Agents

SPADE: Sketch-guided Path Planning Augmented with Diffusion Experts

DGX agent

arXiv:2606.03512v1 Announce Type: cross Abstract: Path planning is essential for Autonomous Mobile Robots (AMRs). Conventional methods for incorporating human preferences into planning typically rely

agentsarxiv-cs-ai
3 Jun 2026
Safety

Taiji: Pareto Optimal Policy Optimization with Semantics-IDs Trade-off for Industrial LLM-Enhanced Recommendation

DGX agent

arXiv:2606.03866v1 Announce Type: cross Abstract: Scaling recommender systems via large language models (LLMs) has become a prominent trend in the industry. However, aligning the LLM's semantic space

safetyarxiv-cs-ai
3 Jun 2026
Research

Testing Most Influential Sets

DGX agent

arXiv:2510.20372v4 Announce Type: replace-cross Abstract: Small influential data subsets can dramatically impact model conclusions, with a few data points overturning key findings. While recent work i

researcharxiv-cs-lg
3 Jun 2026
Agents

The Agent's First Day: Benchmarking Learning, Exploration, and Scheduling in the Workplace Scenarios

DGX agent

arXiv:2601.08173v2 Announce Type: replace Abstract: The rapid evolution of Multi-modal Large Language Models (MLLMs) has advanced workflow automation; however, existing research mainly targets perform

agentsarxiv-cs-ai
3 Jun 2026
Research

The next chapter in flood resilience: Open sourcing Google’s hydrology framework

DGX agent

Google has open-sourced its flood forecasting framework, which replicates operational FloodHub model training settings and reflects methodology described in a 2024 Nature paper for global ungauged flo

researchgoogle-research
3 Jun 2026
Research

The Unsampled Truth: Psychometrics in SLMs Measure Prompt Artifacts, Not Psychological Constructs

DGX agent

arXiv:2606.03357v1 Announce Type: cross Abstract: When prompting SLMs for psychometric assessments, researchers assume the outputs reflect semantic reasoning. We evaluate this premise across 13 open-w

researcharxiv-cs-ai
3 Jun 2026
Applications

Topics as Proxies for Sociodemographics: How Conversational Context Affects LLM Answers

DGX agent

arXiv:2606.02776v1 Announce Type: new Abstract: When large language models (LLMs) are used in high-stakes scenarios, such as legal, medical and financial advice, even a single conversation history is

applicationsarxiv-cs-cl
3 Jun 2026
Agents

Uncertainty-Aware Clarification in LLM Agents with Information Gain

DGX agent

arXiv:2606.03135v1 Announce Type: new Abstract: Large Language Model (LLM) agents often operate under underspecified user instructions, where latent uncertainty over user intent leads to erroneous too

agentsarxiv-cs-ai
3 Jun 2026
Local Ai

v0.30.3

DGX agent

Ollama v0.30.3 adds support for the Gemma 4-12B model . This is a minor patch release that builds on the v0.30 series, which provides improved compatibility and performance improvements. The release w

local-aiollama-releases
3 Jun 2026
Local Ai

Visual Instruction Tuning Aligns Modalities through Abstraction

DGX agent

arXiv:2606.03871v1 Announce Type: cross Abstract: Visual instruction tuning effectively adapts a pre-trained Large Language Model (LLM) to process image information alongside text. Yet, it remains unc

local-aiarxiv-cs-cl
3 Jun 2026
Safety

When Attention Collapses: Stage-Aware Visual Token Pruning from Structure to Semantics

DGX agent

arXiv:2606.03569v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) have demonstrated remarkable capabilities but suffer from significant computational overhead during inference. While vis

safetyarxiv-cs-ai
3 Jun 2026
Local Ai

When RLHF Fails: A Mechanistic Taxonomy of Reward Hacking, Collapse, and Evaluator Gaming

DGX agent

arXiv:2606.03238v1 Announce Type: cross Abstract: Reinforcement learning from human feedback (RLHF) makes large-scale post-training possible by replacing an underspecified human objective with learned

local-aiarxiv-cs-ai
3 Jun 2026
← Previous
1…960961962963964…1292
Next →