AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,428
  • Agents7,398
  • Applications5,301
  • Concepts5
  • Hardware1,785
  • Industry6,113
  • Local Ai4,833
  • Model Releases23,177
  • Research19,713
  • Safety13,092
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,428
  • Agents7,398
  • Applications5,301
  • Concepts5
  • Hardware1,785
  • Industry6,113
  • Local Ai4,833
  • Model Releases23,177
  • Research19,713
  • Safety13,092
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
AllBlog
86,428Total entries
1Added by human
86,427Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,016 results
Tutorials

Learning and Adaptation in Wire Arc Additive Manufacturing Bead Geometry Control

DGX agent

arXiv:2605.29144v1 Announce Type: new Abstract: Robotics Wire Arc Additive Manufacturing (WAAM) is governed by complex and nonlinear process dynamics coupling thermal field to the build geometry. The

tutorialsarxiv-cs-ro
29 May 2026
Research
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Learning Context-Conditioned Predicate Semantics via Prototype Feedback

DGX agent

arXiv:2605.29610v1 Announce Type: cross Abstract: In scene graph generation, a central challenge is modeling polysemous predicates whose meanings shift across contexts. Prior approaches address this i

researcharxiv-cs-ai
29 May 2026
Tutorials

Learning Robust and Task-Invariant Functional Representation from fMRI through Siamese Self-Supervised Learning

DGX agent

arXiv:2605.28990v1 Announce Type: new Abstract: Functional magnetic resonance imaging (fMRI) is a powerful tool for investigating human brain function. However, the high cost of data acquisition and t

tutorialsarxiv-cs-lg
29 May 2026
Research

Lightweight Complementary-Cue Fusion for Robust Video Face Forgery Detection

DGX agent

arXiv:2605.29092v1 Announce Type: new Abstract: Current face video forgery detectors use wide or dual-stream backbones. We show that a single, lightweight fusion of two handcrafted cues can achieve hi

researcharxiv-cs-cv
29 May 2026
Research

Matching Rates and Optimal Allocation for Federated Probe-Logit Distillation under Heterogeneous Bandwidth Budgets

DGX agent

arXiv:2605.29642v1 Announce Type: cross Abstract: In federated language modeling, K nodes each hold n samples but cannot pool data or exchange full-precision gradients or weights. We study the minimax

researcharxiv-cs-lg
29 May 2026
Research

MedCase-Structured: A Text-to-FHIR Dataset for Benchmarking Diagnostic Reasoning in Clinically Realistic EHR Settings

DGX agent

arXiv:2605.30295v1 Announce Type: cross Abstract: Large language models (LLMs) show promise for clinical reasoning and decision support, but evaluation in realistic, electronic health record-congruent

researcharxiv-cs-ai
29 May 2026
Local Ai

MediHive: A Decentralized Agent Collective for Medical Reasoning

DGX agent

arXiv:2603.27150v2 Announce Type: replace Abstract: Large language models (LLMs) have revolutionized medical reasoning tasks, yet single-agent systems often falter on complex, interdisciplinary proble

local-aiarxiv-cs-ai
29 May 2026
Safety

MetaRanker: Human-in-the-loop Active Ranking for Metalens Image Quality

DGX agent

arXiv:2605.29212v1 Announce Type: new Abstract: Image quality in modern imaging systems emerges from the coupled effects of the sensor, optics, and computational reconstruction. Ultra-thin metalenses

safetyarxiv-cs-cv
29 May 2026
Safety

Metric-Dependent Annotation Saturation for Learning from Label Distributions

DGX agent

arXiv:2605.29797v1 Announce Type: new Abstract: When annotators disagree on a label, the disagreement itself carries signal -- and the number of annotators needed to capture it depends on the evaluati

safetyarxiv-cs-cl
29 May 2026
Safety

Mining or Synthesis? Rethinking Exploration Efficiency in Iterative Alignment of Mathematical Reasoning

DGX agent

arXiv:2602.05370v3 Announce Type: replace Abstract: Iterative Direct Preference Optimization (DPO) has emerged as a widely used paradigm for aligning Large Language Models on reasoning tasks. Existing

safetyarxiv-cs-cl
29 May 2026
Applications

MOO: A Multi-view Oriented Observations Dataset for Viewpoint Analysis in Cattle Re-Identification

DGX agent

arXiv:2603.04314v2 Announce Type: replace-cross Abstract: Animal re-identification (ReID) faces critical challenges due to viewpoint variations, particularly in Aerial-Ground (AG-ReID) settings where

applicationsarxiv-cs-ai
29 May 2026
Agents

MOOSE-Copilot: A Web-Based Interactive Assistant for Unified Exploratory and Fine-Grained Scientific Hypothesis Discovery

DGX agent

arXiv:2605.29475v1 Announce Type: cross Abstract: Large language models (LLMs) show remarkable potential in scientific hypothesis discovery. However, existing approaches face two critical limitations:

agentsarxiv-cs-ai
29 May 2026
Research

Neural-Behavioral Representation of Natural Whole-body Movement in Monkeys

DGX agent

arXiv:2605.29355v1 Announce Type: new Abstract: Understanding how cortical activity represents natural whole-body behaviors in primates remains challenging. Limited by the diversity of movements and i

researcharxiv-cs-lg
29 May 2026
Research

OccamToken: Efficient VLM Inference with Training-Free and Budget-Adaptive Token Pruning

DGX agent

arXiv:2605.29657v1 Announce Type: cross Abstract: Vision-language models (VLMs) rely on long visual token sequences for visual understanding, making the prefill stage expensive in both computation and

researcharxiv-cs-ai
29 May 2026
Tutorials

OmniAID: Decoupling Semantic and Artifacts for Universal AI-Generated Image Detection in the Wild

DGX agent

arXiv:2511.08423v3 Announce Type: replace Abstract: A truly universal AI-Generated Image (AIGI) detector must simultaneously generalize across diverse generative models and varied semantic content. Cu

tutorialsarxiv-cs-cv
29 May 2026
Tutorials

One Mask to Rule Them All: On Hidden Facts after Editing and How to Find Them

DGX agent

arXiv:2605.28839v1 Announce Type: new Abstract: Knowledge editing methods such as ROME and MEMIT update factual associations in transformer models by modifying MLP weights. While evaluated mainly by o

tutorialsarxiv-cs-lg
29 May 2026
Safety

Orthogonal Negative Guidance in Attention Feature Space for Text-to-Image Generation

DGX agent

arXiv:2605.29390v1 Announce Type: new Abstract: Text-to-image (T2I) models have become increasingly capable of generating high-quality images. Yet, enforcing the explicit absence of a specified object

safetyarxiv-cs-cv
29 May 2026
Agents

Parse PDFs at lightspeed (this video is at 1x) Absolute cinema

DGX agent

Parse PDFs at lightspeed (this video is at 1x) Absolute cinema Media We've created the world's fastest PDF parser ⚡️ And it's more accurate than any other open-source, model-free PDF parser out there

agentsjerry-liu--x
29 May 2026
Safety

PersonaAgent: Bridging Memory and Action for Personalized LLM Agents

DGX agent

arXiv:2506.06254v2 Announce Type: replace Abstract: Large Language Model (LLM) empowered agents have recently emerged as advanced paradigms that exhibit impressive capabilities in a wide range of doma

safetyarxiv-cs-ai
29 May 2026
Safety

RL2ML: Finite-Rollout Surrogate Objectives from Reinforcement Learning to Maximum Likelihood

DGX agent

arXiv:2605.30154v1 Announce Type: new Abstract: Correctness-based Reinforcement Learning with Verifiable Rewards (RLVR) trains language models from binary feedback on sampled outputs, but the objectiv

safetyarxiv-cs-lg
29 May 2026
Safety

Rooted Absorbed Prefix Trajectory Balance with Submodular Replay for GFlowNet Training

DGX agent

arXiv:2603.00454v2 Announce Type: replace-cross Abstract: Generative Flow Networks (GFlowNets) enable fine-tuning large language models to approximate reward-proportional posteriors, but they remain p

safetyarxiv-cs-ai
29 May 2026
Hardware

Run Step 3.7 Flash on NVIDIA GPUs with Enterprise-Ready Multimodal AI

DGX agent

Step 3.7 Flash is a 198B-parameter Mixture-of-Experts vision-language model designed for enterprise-scale production workloads, featuring native image and video input, a 256k context window, and confi

hardwarenvidia-developer
29 May 2026
Research

S2MDF: A Plug-And-Play Layer for Intersection-Free Multi-Object Signed Distance Fields

DGX agent

arXiv:2605.29761v1 Announce Type: new Abstract: Compositional implicit surface representations model scenes as collections of objects, each encoded by a Signed Distance Field (SDF). A fundamental limi

researcharxiv-cs-cv
29 May 2026
Safety

SafeRx-Agent: A Knowledge-Grounded Multi-Agent Framework for Safe and Explainable Medication Recommendation

DGX agent

arXiv:2605.29146v1 Announce Type: cross Abstract: Medication recommendation predicts medications for patient visits, but existing methods still face two key challenges. At the model level, traditional

safetyarxiv-cs-ai
29 May 2026
Research

SAVAA: Mitigating Hallucinations in LVLMs via Step-wise Adaptive Visual Attention Amplification

DGX agent

arXiv:2602.13600v2 Announce Type: replace Abstract: A line of recent training-free methods for mitigating hallucinations in large vision-language models (LVLMs) operates by amplifying attention to vis

researcharxiv-cs-cv
29 May 2026
Agents

SEAL: Can Saturated Benchmarks Be Revived by LLM-as-a-Meta-Judge?

DGX agent

arXiv:2605.30104v1 Announce Type: new Abstract: Widely used language-model benchmarks are increasingly saturated, with frontier systems often receiving near-tied scores that standard metrics cannot re

agentsarxiv-cs-cl
29 May 2026
Research

Solving Integer Linear Programming with Parallel Tempering

DGX agent

arXiv:2605.29366v1 Announce Type: new Abstract: Integer Linear Programming (ILP) serves as a versatile framework for modeling a wide range of combinatorial optimization problems, typically addressed b

researcharxiv-cs-lg
29 May 2026
Research

SuperVoxelGPT: Adaptive and Ordered 3D Tokenization for Autoregressive Shape Generation

DGX agent

arXiv:2605.29655v1 Announce Type: new Abstract: Autoregressive multimodal large language models (MLLMs) enable 3D generation but struggle to scale to high-resolution shapes due to inadequate 3D tokeni

researcharxiv-cs-cv
29 May 2026
Safety

Teaching Values to Machines: Simulating Human-Like Behavior in LLMs

DGX agent

arXiv:2605.30036v1 Announce Type: new Abstract: Large Language Models (LLMs) demonstrate a remarkable capacity to adopt different personas and roles; however, it remains unclear whether they can manif

safetyarxiv-cs-ai
29 May 2026
Safety

The Anatomy of Conversational Scams: A Topic-Based Red Teaming Analysis of Multi-Turn Interactions in LLMs

DGX agent

arXiv:2601.03134v2 Announce Type: replace Abstract: As LLMs gain persuasive capabilities through extended dialogues, they create new opportunities for studying adversarial conversational behavior in e

safetyarxiv-cs-cl
29 May 2026
Agents

Towards Verifiable Multimodal Deep Research: A Multi-Agent Harness for Interleaved Report Generation

DGX agent

arXiv:2605.29861v1 Announce Type: cross Abstract: Large Language Models (LLMs) have advanced autonomous agents from deep search, which retrieves concise factual answers, to deep research, which synthe

agentsarxiv-cs-ai
29 May 2026
Research

Turbulence-Robust Dynamic Object Segmentation with Multi-Signal Priors and SAM2 Refinement

DGX agent

arXiv:2605.29292v1 Announce Type: new Abstract: This technical report presents our solution for the CVPR 2026 UG2+ Challenge Track 3: Dynamic Object Segmentation in Turbulence (DOST). We design a trai

researcharxiv-cs-cv
29 May 2026
Local Ai

UI-KOBE: Knowledge-Oriented Behavior Exploration for Lightweight Graph-Guided GUI Agents

DGX agent

arXiv:2605.29534v1 Announce Type: new Abstract: Recent advances in mobile GUI agents have shown strong potential for automating mobile tasks, but most effective systems still depend on large vision-la

local-aiarxiv-cs-ai
29 May 2026
Research

Understanding the Ability of LLMs to Handle Character-Level Perturbation

DGX agent

arXiv:2510.14365v4 Announce Type: replace Abstract: This work investigates the resilience of contemporary large language models (LLMs) against frequent character-level perturbations. We examine three

researcharxiv-cs-cl
29 May 2026
Tutorials

Uni-RCM: Unified Reference-guided Cross-modal Mapping for Multi-Class Anomaly Detection

DGX agent

arXiv:2605.29455v1 Announce Type: new Abstract: Multi-modal industrial anomaly detection typically relies on separate models for each product category, fundamentally limiting practical scalability. Wh

tutorialsarxiv-cs-cv
29 May 2026
Local Ai

Unifying Temporal and Structural Credit Assignment in LLM-Based Multi-Agent Prompt Optimization

DGX agent

arXiv:2605.30227v1 Announce Type: cross Abstract: While Multi-Agent Systems (MAS) empower Large Language Models to tackle complex reasoning tasks through collaborative interaction, optimizing their dy

local-aiarxiv-cs-ai
29 May 2026
Safety

User-Aware Active Knowledge Acquisition for Emotional Support Dialogue

DGX agent

arXiv:2605.29715v1 Announce Type: new Abstract: Emotional support plays an important role in dialogue systems, and its success depends on adapting to a user's evolving and implicit needs across multi-

safetyarxiv-cs-cl
29 May 2026
Safety

V2XCrafter: Learning to Generate Driving Scene Across Agents

DGX agent

arXiv:2605.29471v1 Announce Type: new Abstract: Collaborative driving systems leverage vehicle-to-everything (V2X) communication for multi-agent collaborative perception to enhance driving safety, yet

safetyarxiv-cs-cv
29 May 2026
Local Ai

Video fusion/join options

DGX agent

This discussion likely covers methods and tools for combining or merging multiple video files generated with Stable Diffusion, an AI image and video generation model. Community members likely share re

local-air-stablediffusion
29 May 2026
Agents

VikingMem: A Memory Base Management System for Stateful LLM-based Applications

DGX agent

arXiv:2605.29640v1 Announce Type: new Abstract: Large Language Models have revolutionized interactive applications; however, their finite context windows pose a critical data management challenge for

agentsarxiv-cs-ai
29 May 2026
Local Ai

[Web Chat para Ollama: Tu propia IA local, gratis y con privacidad total] [Local AI Web Chat for Ollama: 100% private, free, and runs entirely on your machine]

DGX agent

Web Chat para Ollama is a tool that enables users to run a private, free AI chat interface locally on their own machines using Ollama, an open-source framework for running large language models. The p

local-air-ollama
29 May 2026
Hardware

What is the best used or refurbished laptop with GPU for open source Imege generation?

DGX agent

This Reddit discussion in the StableDiffusion community addresses recommendations for affordable, used or refurbished laptops equipped with GPUs suitable for running open-source image generation model

hardwarer-stablediffusion
29 May 2026
Applications

When Does Persona Prompting Actually Help? A Retrieval and Metric Analysis of Expert Role Injection in LLMs

DGX agent

arXiv:2605.29420v1 Announce Type: new Abstract: Persona prompting is widely used to steer large language models, yet its practical value remains unclear. Prior work often evaluates persona prompting u

applicationsarxiv-cs-ai
29 May 2026
Agents

WorldMemArena: Evaluating Multimodal Agent Memory Through Action-World Interaction

DGX agent

arXiv:2605.29341v1 Announce Type: cross Abstract: Multimodal large language models are increasingly deployed as long-horizon agents, where memory must do more than recall: it must track an evolving wo

agentsarxiv-cs-cl
29 May 2026
Research

X-GS: An Extensible Framework for Perceiving and Thinking via 3D Gaussian Splatting

DGX agent

arXiv:2603.09632v3 Announce Type: replace-cross Abstract: 3D Gaussian Splatting (3DGS) has emerged as a powerful technique for novel view synthesis, subsequently extending into numerous spatial AI app

researcharxiv-cs-cl
29 May 2026
Research

A New Era of Innovation: Google Research at I/O 2026

DGX agent

Google Research at I/O 2026 showcased new Gemini AI models including Gemini Omni, which can create content from any input starting with video, and Gemini 3.5 Flash, combining frontier intelligence wit

researchgoogle-research
28 May 2026
Hardware

A Unified Structured Query Understanding Framework for Industrial Semantic Search

DGX agent

arXiv:2605.27441v1 Announce Type: cross Abstract: Query understanding in large-scale industrial search systems is typically implemented as a cascade of disparate, task-specific components. While indiv

hardwarearxiv-cs-lg
28 May 2026
Research

ABot-OCR Technical Report

DGX agent

arXiv:2605.27978v1 Announce Type: new Abstract: We introduce ABot-OCR, an end-to-end vision-language model that transcribes a page image directly into clean Markdown in a single forward pass. By doing

researcharxiv-cs-cv
28 May 2026
← Previous
1…967968969970971…1292
Next →