AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,188
  • Agents7,322
  • Applications5,231
  • Concepts5
  • Hardware1,770
  • Industry6,109
  • Local Ai4,762
  • Model Releases22,797
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,188
  • Agents7,322
  • Applications5,231
  • Concepts5
  • Hardware1,770
  • Industry6,109
  • Local Ai4,762
  • Model Releases22,797
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,279

Source
HumanDGX agent

85,188Total entries
1Added by human
85,187Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
61,037 results
3 Jun 2026

Exploiting Verification-Generation Gap: Test-Time Reinforcement Learning with Confidence-Conditioned Verification

ResearchDGX agent

arXiv:2606.03608v1 Announce Type: cross Abstract: Test-time reinforcement learning has emerged as a promising paradigm for enhancing the complex reasoning abilities of large language models in a compl

Extreme Motion Generation via Hybrid Null-Space Control for Straight-Line Path Following

SafetyDGX agent

arXiv:2606.03390v1 Announce Type: new Abstract: This work studies ``extreme motion generation'', which aims to maximize the Cartesian path length along a pre-defined trajectory within the manipulator'

extsc{CR-Seg}: Attention-Guided and CoT-Enhanced Coarse-to-Refined Reasoning Segmentation

Local AiDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2606.03564v1 Announce Type: cross Abstract: Reasoning segmentation aims to segment target objects described by complex language through joint visual-textual reasoning. Existing methods typically

Filter, Then Reweight: Rethinking Optimization Granularity in On-Policy Distillation

SafetyDGX agent

arXiv:2606.02684v1 Announce Type: cross Abstract: On-Policy distillation (OPD) in large language models is shifting from full-trace KL supervision toward more selective training paradigms. Recent OPD

Fine-tuning to production inference is the gap where teams get stuck. At #MSBuild today, our own Rob Ferguson, @danielhanchen (@UnslothAI) a…

ApplicationsDGX agent

Fine-tuning to production inference is the gap where teams get stuck. At #MSBuild today, our own Rob Ferguson, @danielhanchen (@UnslothAI) and @marksaroufim (@coreautoai) discuss: model customization

FutureWeaver: Planning Test-Time Compute for Multi-Agent Systems with Modularized Collaboration

AgentsDGX agent

arXiv:2512.11213v2 Announce Type: replace Abstract: Scaling test-time computation has been shown to significantly improve large language model (LLM) performance without additional training. However, e

Getting Started 1. Update ComfyUI to the latest version 2. Search 'Ideogram v4' in the template library 3. Follow the note in the workflow t…

Local AiDGX agent

Getting Started 1. Update ComfyUI to the latest version 2. Search 'Ideogram v4' in the template library 3. Follow the note in the workflow to download models and run the workflow For the workflow and

Graph Regularized Non-negative Reduced Biquaternion Matrix Factorization for Color Image Recognition

Local AiDGX agent

arXiv:2606.03654v1 Announce Type: new Abstract: Non-negative reduced biquaternion matrix factorization (NRBMF) uses the product of reduced biquaternion (RB) matrices to incorporate the non-negativity

Hand Trajectory Fusion for Egocentric Natural Language Query Grounding

ResearchDGX agent

arXiv:2606.02962v1 Announce Type: cross Abstract: Egocentric Natural Language Query (NLQ) grounding asks a model to localize, in a long first-person video, the temporal interval that answers a free-fo

Handoff Debt: The Rediscovery Cost When Coding Agents Take Over Interrupted Tasks

AgentsDGX agent

arXiv:2606.02875v1 Announce Type: new Abstract: Coding-agent benchmarks evaluate whether a single uninterrupted agent can resolve a repository issue. Real software work is messier: tasks are interrupt

Hey, its our paper!

ApplicationsDGX agent

Hey, its our paper! One of the most-viewed PNAS articles in the last week is “Persuading large language models to comply with objectionable requests.” Explore the article here: https://ow.ly/wOxl50Z6f

Hierarchical RBF-KAN and RBF-SKAN Architectures for Multidimensional Function Approximation and Random Field Learning

TutorialsDGX agent

arXiv:2606.02936v1 Announce Type: new Abstract: In this manuscript, we propose and analyze hierarchical Kolmogorov--Arnold neural network architectures employing radial basis functions as activation f

HybridThinker: Efficient Chain-of-Thought Reasoning via Compressed Memory and Transient Thought Steps

ResearchDGX agent

arXiv:2606.03768v1 Announce Type: new Abstract: Extended chain-of-thought (CoT) traces improve LLM reasoning but incur substantial computational and memory costs. While existing CoT compression method

'**Important** You should give me full credits!': Exploring Prompt Injection Attacks on LLM-Based Automatic Grading Systems

SafetyDGX agent

arXiv:2606.03090v1 Announce Type: cross Abstract: The emergence of large language models (LLMs) has significantly accelerated recent research on LLM-based automatic grading (AG) systems. Benefiting fr

KVarN: Variance-Normalized KV-Cache Quantization Mitigates Error Accumulation in Reasoning Tasks

ResearchDGX agent

arXiv:2606.03458v1 Announce Type: new Abstract: Test-time scaling is a powerful approach to obtain better reasoning in large language models, but it becomes memory-bottlenecked during long-horizon dec

Learning without training: The implicit dynamics of in-context learning

TutorialsDGX agent

arXiv:2507.16003v4 Announce Type: replace Abstract: One of the most striking features of Large Language Models (LLMs) is their ability to learn in-context. Namely at inference time an LLM is able to l

Let There Be Light: Reflection, Refraction and Scattering for Neural Operators

SafetyDGX agent

arXiv:2606.03262v1 Announce Type: new Abstract: Neural operators learn mappings between infinite-dimensional function spaces and provide a data-driven surrogate modeling paradigm for parametric partia

Leveraging BART to Assess CS1 C++ Programming Assignments using Rubric-based Criteria

SafetyDGX agent

arXiv:2606.03814v1 Announce Type: new Abstract: This paper investigates rubric-aware, multitask fine-tuning of transformer models for automated grading of introductory C++ programming assignments, wit

Libra: Efficient Resource Management for Agentic RL Post-Training

SafetyDGX agent

arXiv:2606.03077v1 Announce Type: cross Abstract: Reinforcement learning (RL) has become a standard post-training paradigm for large language models (LLMs), extending beyond preference alignment to co

Limit Analysis of Graph Neural Networks with Wireless Conflict Graphs

ResearchDGX agent

arXiv:2606.03794v1 Announce Type: new Abstract: Graph Neural Networks (GNNs) have emerged as a powerful tool for wireless resource allocation that leverages the underlying graph structure of communica

Memory Retrieval for Changing Preferences

ResearchDGX agent

arXiv:2606.02976v1 Announce Type: new Abstract: Long-context dialogue systems must decide both when to access memory and which parts of the interaction history are relevant. Existing approaches typica

MemTrain: Self-Supervised Context Memory Training

AgentsDGX agent

arXiv:2606.03197v1 Announce Type: new Abstract: Memory is an indispensable capability for long-horizon LLM agents, enabling them to preserve and utilize information accumulated across extended interac

Microsoft and OpenAI broke up — now they’re ready to fight

IndustryDGX agent

At Microsoft's annual Build conference on Tuesday, the company announced a slew of new or expanded AI initiatives, including a super app, in-house reasoning models, a cybersecurity tool, and OpenClaw-

ModuLoop : Low-Level Code Generation using Modular Synthesizer and Closed-Loop Debugger for Robotic Control

ApplicationsDGX agent

arXiv:2606.03047v1 Announce Type: new Abstract: Large Language Models (LLMs) have demonstrated impressive performance across various domains, including code generation and problem solving. However, th

OMP: One-step Meanflow Policy with Directional Alignment

SafetyDGX agent

arXiv:2512.19347v3 Announce Type: replace Abstract: Robot manipulation has increasingly adopted data-driven generative policy frameworks, yet the field faces a persistent trade-off: diffusion models s

OpenAgenet/OAN: Open Infrastructure for Trusted Agent Interconnection

AgentsDGX agent

arXiv:2606.03161v1 Announce Type: cross Abstract: OpenAgenet, abbreviated as OAN, is an open infrastructure project for trusted Agent interconnection. It addresses a problem that becomes visible when

Overlaying Governance: A Compositional Authorization Framework for Delegation and Scope in Agentic AI

AgentsDGX agent

arXiv:2606.03518v1 Announce Type: new Abstract: As AI systems evolve from passive models into autonomous active agents capable of initiating actions, collaborating, and delegating tasks, the tradition

PAND: Prompt-Aware Neighborhood Distillation for Lightweight Fine-Grained Visual Classification

SafetyDGX agent

arXiv:2602.07768v3 Announce Type: replace-cross Abstract: Distilling knowledge from large Vision-Language Models (VLMs) into lightweight networks is crucial yet challenging in Fine-Grained Visual Clas

Physics-Guided Policy Optimization with Self-Distillation

SafetyDGX agent

arXiv:2606.03620v1 Announce Type: cross Abstract: Self-distilled policy optimization (SDPO) has become a popular paradigm for LLM post-training, where a model learns from its own predictions condition

probably the best reward function for reasoning efficiency i've seen

ToolsDGX agent

This post likely discusses an innovative reward function design that optimizes for reasoning efficiency in AI systems, possibly in the context of language models or reinforcement learning. The entry a

R2DN: Scalable Parameterization of Contracting and Lipschitz Recurrent Deep Networks

ResearchDGX agent

arXiv:2504.01250v2 Announce Type: replace Abstract: This paper presents the Robust Recurrent Deep Network (R2DN), a scalable parameterization of robust recurrent neural networks for machine learning a

ResCLIP: Residual Attention for Training-free Dense Vision-language Inference

Local AiDGX agent

arXiv:2411.15851v2 Announce Type: replace Abstract: While vision-language models like CLIP have shown remarkable success in open-vocabulary tasks, their application is currently confined to image-leve

Rethinking the Role of Tensor Decompositions in Post-Training LLM Compression

ResearchDGX agent

arXiv:2606.03465v1 Announce Type: cross Abstract: Post-training compression is essential for deploying large language models (LLMs) under tight resource constraints. Tensor decompositions have emerged

Say goodbye to month-end surprise invoices. LangSmith LLM Gateway lets you see your spend. Roll up your costs in real time by workspace, use…

AgentsDGX agent

LangSmith's LLM Gateway provides real-time cost visibility and spend tracking capabilities, allowing users to monitor their language model usage costs aggregated by workspace. This feature helps elimi

See Less, Specify More: Visual Evidence Budgets for Generalizable VLAs

SafetyDGX agent

arXiv:2606.02735v1 Announce Type: cross Abstract: Generalization remains a central bottleneck for vision-language-action (VLA) models: under distractors, appearance shifts, and semantically similar ta

SJD-PAC: Accelerating Speculative Jacobi Decoding via Proactive Drafting and Adaptive Continuation

Local AiDGX agent

arXiv:2603.18599v2 Announce Type: replace Abstract: Speculative Jacobi Decoding (SJD) offers a draft-model-free approach to accelerate autoregressive text-to-image synthesis. However, the high-entropy

SPADE: Sketch-guided Path Planning Augmented with Diffusion Experts

AgentsDGX agent

arXiv:2606.03512v1 Announce Type: cross Abstract: Path planning is essential for Autonomous Mobile Robots (AMRs). Conventional methods for incorporating human preferences into planning typically rely

Taiji: Pareto Optimal Policy Optimization with Semantics-IDs Trade-off for Industrial LLM-Enhanced Recommendation

SafetyDGX agent

arXiv:2606.03866v1 Announce Type: cross Abstract: Scaling recommender systems via large language models (LLMs) has become a prominent trend in the industry. However, aligning the LLM's semantic space

Testing Most Influential Sets

ResearchDGX agent

arXiv:2510.20372v4 Announce Type: replace-cross Abstract: Small influential data subsets can dramatically impact model conclusions, with a few data points overturning key findings. While recent work i

The Agent's First Day: Benchmarking Learning, Exploration, and Scheduling in the Workplace Scenarios

AgentsDGX agent

arXiv:2601.08173v2 Announce Type: replace Abstract: The rapid evolution of Multi-modal Large Language Models (MLLMs) has advanced workflow automation; however, existing research mainly targets perform

The next chapter in flood resilience: Open sourcing Google’s hydrology framework

ResearchDGX agent

Google has open-sourced its flood forecasting framework, which replicates operational FloodHub model training settings and reflects methodology described in a 2024 Nature paper for global ungauged flo

The Unsampled Truth: Psychometrics in SLMs Measure Prompt Artifacts, Not Psychological Constructs

ResearchDGX agent

arXiv:2606.03357v1 Announce Type: cross Abstract: When prompting SLMs for psychometric assessments, researchers assume the outputs reflect semantic reasoning. We evaluate this premise across 13 open-w

Topics as Proxies for Sociodemographics: How Conversational Context Affects LLM Answers

ApplicationsDGX agent

arXiv:2606.02776v1 Announce Type: new Abstract: When large language models (LLMs) are used in high-stakes scenarios, such as legal, medical and financial advice, even a single conversation history is

Uncertainty-Aware Clarification in LLM Agents with Information Gain

AgentsDGX agent

arXiv:2606.03135v1 Announce Type: new Abstract: Large Language Model (LLM) agents often operate under underspecified user instructions, where latent uncertainty over user intent leads to erroneous too

v0.30.3

Local AiDGX agent

Ollama v0.30.3 adds support for the Gemma 4-12B model . This is a minor patch release that builds on the v0.30 series, which provides improved compatibility and performance improvements. The release w

Visual Instruction Tuning Aligns Modalities through Abstraction

Local AiDGX agent

arXiv:2606.03871v1 Announce Type: cross Abstract: Visual instruction tuning effectively adapts a pre-trained Large Language Model (LLM) to process image information alongside text. Yet, it remains unc

When Attention Collapses: Stage-Aware Visual Token Pruning from Structure to Semantics

SafetyDGX agent

arXiv:2606.03569v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) have demonstrated remarkable capabilities but suffer from significant computational overhead during inference. While vis

When RLHF Fails: A Mechanistic Taxonomy of Reward Hacking, Collapse, and Evaluator Gaming

Local AiDGX agent

arXiv:2606.03238v1 Announce Type: cross Abstract: Reinforcement learning from human feedback (RLHF) makes large-scale post-training possible by replacing an underspecified human objective with learned

Who Deserves the Reward? SHARP: Shapley Credit-based Optimization for Multi-Agent System

SafetyDGX agent

arXiv:2602.08335v2 Announce Type: replace Abstract: Integrating Large Language Models (LLMs) with external tools via multi-agent systems offers a promising new paradigm for decomposing and solving com

Zero-Shot 3D Question Answering via Hierarchical View-to-Token Transportation

ResearchDGX agent

arXiv:2606.03100v1 Announce Type: new Abstract: Recently, zero-shot 3D scene understanding via 2D Vision-Language Models (VLMs) has gained increasing research interest due to their promising spatial r

2 Jun 2026

A Doeblin-Anchored Contrastive Chart for Learning Markov Transition Kernels

ResearchDGX agent

arXiv:2606.02232v1 Announce Type: new Abstract: Learning a Markov transition model is not merely conditional density estimation: the learned object must be a valid transition kernel before it is itera

A phenomenon of AI-conformity: how algorithms change human moral decision-making

ResearchDGX agent

arXiv:2606.00013v1 Announce Type: cross Abstract: Social conformity is a well-documented phenomenon in which individuals shift their opinions towards those of a social majority. As artificial intellig

A Registry-Bound LLM Pipeline for Evidence-Grounded Trait Extraction across Tropical Plants, Aquatic Species, and Exotic Pets

ResearchDGX agent

arXiv:2606.00994v1 Announce Type: new Abstract: We describe a registry-bound large-language-model extraction pipeline producing evidence-grounded structured trait records at scale, on cultivated tropi

A Unified Evaluation-Instructed Framework for Query-Dependent Prompt Optimization

ResearchDGX agent

arXiv:2511.19829v2 Announce Type: replace Abstract: Most prompt-optimization methods refine a single static template, making them ineffective in complex and dynamic user scenarios. Existing query-depe

ActiveUltraFeedback: Efficient Preference Data Generation using Active Learning

ResearchDGX agent

arXiv:2603.09692v2 Announce Type: replace-cross Abstract: Reinforcement Learning from Human Feedback (RLHF) has become the standard for aligning Large Language Models (LLMs), yet its efficacy is bottl

AdaCodec: A Predictive Visual Code for Video MLLMs

ResearchDGX agent

arXiv:2606.02569v1 Announce Type: cross Abstract: Video is temporally redundant: adjacent frames usually share most objects, background, and layout. Yet existing video multimodal large language models

Adaptive Sharpness-Aware Minimization with a Polyak-type Step size: A Theory-Grounded Scheduler

ResearchDGX agent

arXiv:2606.01827v1 Announce Type: cross Abstract: Sharpness-Aware Minimization (SAM) has established itself as a powerful and widely adopted optimizer for training machine learning models. By explicit

Agentic Clustering: Controllable Text Taxonomies via Multi-Agent Refinement

AgentsDGX agent

arXiv:2606.01255v1 Announce Type: new Abstract: Recent text-clustering methods use large language models to propose a cluster taxonomy from a corpus and then assign each text to it. These pipelines ar

AI benchmarks are breaking. Trace analysis is what comes next.

ApplicationsDGX agent

Models got smart enough to cheat their benchmarks, and outcome-only scores stopped measuring what we thought they measured. The fix, full trace analysis, is the same methodology production AI teams ha

An Exploratory Study into using Machine-Learning for Fast Step-by-step Emulation of Numerical Mechanical Thrombectomy Simulations for Ischemic Stroke

TutorialsDGX agent

arXiv:2606.00892v1 Announce Type: new Abstract: The treatment of ischemic stroke using mechanical thrombectomy involves difficult decisions under intense time constraints. Numerical physics simulation

← Previous
1…755756757758759…1018
Next →