AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,376
  • Agents7,554
  • Applications5,409
  • Concepts5
  • Hardware1,835
  • Industry6,170
  • Local Ai4,930
  • Model Releases23,883
  • Research20,124
  • Safety13,369
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,376
  • Agents7,554
  • Applications5,409
  • Concepts5
  • Hardware1,835
  • Industry6,170
  • Local Ai4,930
  • Model Releases23,883
  • Research20,124
  • Safety13,369
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

88,376Total entries
1Added by human
88,375Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,602 results
18 May 2026

Beyond Forgetting: Machine Unlearning Elicits Controllable Side Behaviors and Capabilities

ResearchDGX agent

arXiv:2601.21702v3 Announce Type: replace-cross Abstract: We consider Representation Misdirection (RM), a class of large language model (LLM) unlearning methods that achieve forgetting by redirecting

Beyond Sunk Costs: Boosting LLM Pre-training Efficiency via Orthogonal Growth of Mixture-of-Experts

ResearchDGX agent

arXiv:2510.08008v2 Announce Type: replace Abstract: As the computational demands for pre-training Large Language Models (LLMs) continue to surge, the need for efficient training paradigms becomes crit

ChronoEarth-492K: A Large Scale and Long Horizon Spatiotemporal Hyperspectral Earth Observation Dataset and Benchmark

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.15666v1 Announce Type: new Abstract: Hyperspectral imaging (HSI) provides dense spectral information for the Earth's surface, enabling material-level understanding of land cover and ecosyst

CIS-BWE: Chaos-Informed Speech Bandwidth Extension

Model ReleasesDGX agent

arXiv:2507.15970v3 Announce Type: replace-cross Abstract: Recovering high-frequency components lost to bandwidth constraints is crucial for applications ranging from telecommunications to high-fidelit

Contexting as Recommendation: Evolutionary Collaborative Filtering for Context Engineering

TutorialsDGX agent

arXiv:2605.15721v1 Announce Type: new Abstract: Large Language Models (LLMs) are highly sensitive to their input contexts, motivating the development of automated context engineering. However, existin

CryptoBench: A Dynamic Benchmark for Expert-Level Evaluation of LLM Agents in Cryptocurrency

Model ReleasesDGX agent

arXiv:2512.00417v5 Announce Type: replace Abstract: This paper introduces CryptoBench, the first expert-curated, dynamic benchmark designed to rigorously evaluate the real-world capabilities of Large

DeltaPrompts: Escaping the Zero-Delta Trap in Multimodal Distillation

SafetyDGX agent

arXiv:2605.15532v1 Announce Type: cross Abstract: Distillation enables compact Vision-Language Models (VLMs) to obtain strong reasoning capabilities, yet the prompts driving this process are typically

Does Theory of Mind Improvement Really Benefit Human-AI Interactions? Empirical Findings from Interactive Evaluations

ApplicationsDGX agent

arXiv:2605.15205v1 Announce Type: new Abstract: Improving the Theory of Mind (ToM) capability of Large Language Models (LLMs) is crucial for effective social interactions between these AI models and h

DreamSR: Towards Ultra-High-Resolution Image Super-Resolution via a Receptive-Field Enhanced Diffusion Transformer

Local AiDGX agent

arXiv:2605.15682v1 Announce Type: new Abstract: Large-scale pre-trained diffusion models have been extensively adopted for real-world image Super-Resolution because of their powerful generative priors

Dynamic-TreeRPO: Breaking the Independent Trajectory Bottleneck with Structured Sampling

SafetyDGX agent

arXiv:2509.23352v3 Announce Type: replace-cross Abstract: The integration of Reinforcement Learning (RL) into flow matching models for text-to-image (T2I) generation has driven substantial advances in

EduVQA: Towards Concept-Aware Assessment of Educational AI-Generated Videos

Model ReleasesDGX agent

arXiv:2603.03066v2 Announce Type: replace Abstract: Existing AI-generated video quality assessment (AIGVQA) methods mainly focus on global perceptual realism and coarse text-video alignment, while ove

Effective Harness Engineering for Algorithm Discovery with Coding Agents

ResearchDGX agent

arXiv:2605.15221v1 Announce Type: cross Abstract: AlphaEvolve and FunSearch have demonstrated the potential of combining large language models (LLMs) with evolutionary search for automated algorithm d

Fine-Tuning NVIDIA Cosmos Predict 2.5 with LoRA/DoRA for Robot Video Generation

HardwareDGX agent

This guide demonstrates how to fine-tune NVIDIA's Cosmos Predict 2.5 video generation model using parameter-efficient techniques like LoRA (Low-Rank Adaptation) and DoRA (Mixture of Experts-based adap

Flowette: Flow Matching with Graphette Priors for Graph Generation

SafetyDGX agent

arXiv:2602.23566v2 Announce Type: replace-cross Abstract: We study generative modeling of graphs with recurring subgraph motifs. We propose Flowette, a continuous flow matching framework that employs

FRWKV+: Adaptive Periodic-Position Branch Interaction for Frequency-Space Linear Time Series Forecasting

Model ReleasesDGX agent

arXiv:2605.15690v1 Announce Type: new Abstract: Long-term time series forecasting is essential for decision making in energy, finance, transportation, and healthcare systems. Recent lightweight foreca

IO-SVD: Input-Output Whitened SVD for Adaptive-Rank LLM Compression

ResearchDGX agent

arXiv:2605.15626v1 Announce Type: new Abstract: Large language models deliver strong performance across language and reasoning tasks, but their storage and compute costs remain major barriers to deplo

Margin-Adaptive Confidence Ranking for Reliable LLM Judgement

ResearchDGX agent

arXiv:2605.15416v1 Announce Type: cross Abstract: Jung et al. (2025) introduce a hypothesis testing framework for guaranteeing agreement between large language models (LLMs) and human judgments, relyi

MI-CXR: A Benchmark for Longitudinal Reasoning over Multi-Interval Chest X-rays

Model ReleasesDGX agent

arXiv:2605.15574v1 Announce Type: new Abstract: Longitudinal chest X-ray (CXR) interpretation requires reasoning over disease evolution across multiple patient visits, yet most existing medical VQA be

MuteBench: Modality Unavailability Tolerance Evaluation for Incomplete Multimodal Fusion

Model ReleasesDGX agent

arXiv:2605.15235v1 Announce Type: new Abstract: Multimodal physiological data powers clinical AI systems from intensive care units to wearable devices, but sensors routinely fail in practice. Two fail

Not All Tasks Quantize Equally: Fisher-Guided Quantization for Visual Geometry Transformer

Local AiDGX agent

arXiv:2605.15828v1 Announce Type: new Abstract: Feed-forward 3D reconstruction models, represented by Visual Geometry Grounded Transformer (VGGT), jointly predict multiple visual geometry tasks such a

Offline Semantic Guidance for Efficient Vision-Language-Action Policy Distillation

Model ReleasesDGX agent

arXiv:2605.16241v1 Announce Type: cross Abstract: Billion-parameter Vision-Language-Action (VLA) policies have recently shown impressive performance in robotic manipulation, yet their size and inferen

Position: Early-Stage Quality Assurance in Annotation Pipelines Is More Cost-Effective Than Late-Stage Validation

Model ReleasesDGX agent

arXiv:2605.15714v1 Announce Type: cross Abstract: This position paper argues that the machine learning community should prioritize early-stage quality assurance in annotation pipelines over the prevai

Probabilistic Dating of Historical Manuscripts via Evidential Deep Regression on Visual Script Features

Model ReleasesDGX agent

arXiv:2605.06475v1 Announce Type: cross Abstract: We introduce a probabilistic approach for dating historical manuscript pages from visual features alone. Instead of aggregating centuries into classes

RAR: Retrieving And Ranking Augmented MLLMs for Visual Recognition

Model ReleasesDGX agent

arXiv:2403.13805v2 Announce Type: replace-cross Abstract: CLIP (Contrastive Language-Image Pre-training) uses contrastive learning from noise image-text pairs to excel at recognizing a wide array of c

Runtime-Orchestrated Second-Order Optimization for Scalable LLM Training

Model ReleasesDGX agent

arXiv:2605.16184v1 Announce Type: cross Abstract: Second-order methods offer an attractive path toward more sample-efficient LLM training, but their practical use is often blocked by the systems cost

SemanticOpt: Towards LLM-Based Semantic Black-Box Optimization

Model ReleasesDGX agent

arXiv:2510.25404v3 Announce Type: replace-cross Abstract: Optimizing an experimental system can be extremely challenging when each experiment is expensive, time-consuming, or difficult to perform. Exi

Skyra: AI-Generated Video Detection via Grounded Artifact Reasoning

Model ReleasesDGX agent

arXiv:2512.15693v2 Announce Type: replace Abstract: The misuse of AI-driven video generation technologies has raised serious social concerns, highlighting the urgent need for reliable AI-generated vid

STAR: A Stage-attributed Triage and Repair framework for RCA Agents in Microservices

Model ReleasesDGX agent

arXiv:2605.15581v1 Announce Type: new Abstract: LLM-based root cause analysis (RCA) agents have recently emerged as a promising paradigm for incident diagnosis in microservice AIOps. However, their re

T2T-LA: A Topology-to-Topology LLM Agent for Graph Learning with Neither Feature Access nor Task Knowledge

Model ReleasesDGX agent

arXiv:2512.08964v4 Announce Type: replace Abstract: Graph learning aims to convert data into graph representations, which are fundamental to many problems in machine learning for CAD, where circuits,

Variational Autoregressive Networks with probability priors

ResearchDGX agent

arXiv:2605.16020v1 Announce Type: new Abstract: Monte Carlo methods are essential across diverse scientific fields, yet their efficiency is frequently hampered by critical slowing down-a sharp increas

VLMs Trace Without Tracking: Diagnosing Failures in Visual Path Following

Local AiDGX agent

arXiv:2605.15672v1 Announce Type: cross Abstract: Vision-language models (VLMs) achieve strong performance on multimodal benchmarks, but may still lack robust control over basic visual operations. We

VSPO: Vector-Steered Policy Optimization for Behavioral Control

SafetyDGX agent

arXiv:2605.15604v1 Announce Type: cross Abstract: Modern language models often need to optimize a primary accuracy objective while also accommodating secondary behavioral preferences, such as verbosit

Weight Concentration Regularization for Improving Pruning Robustness Under High Sparsity

Model ReleasesDGX agent

arXiv:2511.14282v2 Announce Type: replace-cross Abstract: Deep neural networks achieve outstanding performance across vision and language tasks, yet their large parameter counts limit deployment in re

17 May 2026

Harness profiles! https://docs.langchain.com/oss/python/deepagents/profiles

Model ReleasesDGX agent

Harness profiles! https://docs.langchain.com/oss/python/deepagents/profiles I like what Langchain recently released in their deepagents harness, which is an adapter to modify the syntax of primitive f

Simulate real-world places with Project Genie and Street View

ApplicationsDGX agent

Google DeepMind has integrated its generative world model Project Genie with Street View imagery, allowing AI models to create interactive virtual environments anchored to real-world locations for age

16 May 2026

Real AGI would not do this. Even after a trillion dollars in LLMs still do.

SafetyDGX agent

Real AGI would not do this. Even after a trillion dollars in LLMs still do. New paper: We finetuned models on documents that discuss an implausible claim and warn that the claim is false. Models ended

15 May 2026

A Heterogeneous Temporal Memory Governance Framework for Long-Term LLM Persona Consistency

ResearchDGX agent

arXiv:2605.14802v1 Announce Type: new Abstract: Large language models often suffer from fact loss, timeline confusion, persona drift, and reduced stability during long-range interaction, especially un

A Tutorial on Cognitive Biases in Agentic AI-Driven 6G Autonomous Networks

Model ReleasesDGX agent

arXiv:2510.19973v4 Announce Type: replace-cross Abstract: The path to higher network autonomy in 6G lies beyond the mere optimization of key performance indicators (KPIs), requiring systems that perce

Addressing Terminal Constraints in Data-Driven Demand Response Scheduling

Model ReleasesDGX agent

arXiv:2605.14741v1 Announce Type: cross Abstract: Electrified chemical processes are incentivized by exposure to time-varying electricity markets to operate flexibly, but participating in demand respo

Agentic Design of Compositional Descriptors via Autoresearch for Materials Science Applications

Model ReleasesDGX agent

arXiv:2605.14671v1 Announce Type: cross Abstract: Autoresearch offers a flexible paradigm for automating scientific tasks, in which an AI agent proposes, implements, evaluates, and refines candidate s

Architecture-Aware Explanation Auditing for Industrial Visual Inspection

ResearchDGX agent

arXiv:2605.14255v1 Announce Type: cross Abstract: Industrial visual inspection systems increasingly rely on deep classifiers whose heatmap explanations may appear visually plausible while failing to i

Are Agents Ready to Teach? A Multi-Stage Benchmark for Real-World Teaching Workflows

Model ReleasesDGX agent

arXiv:2605.14322v1 Announce Type: new Abstract: Language agents are increasingly deployed in complex professional workflows, with tutoring emerging as a particularly high-stakes capability that remain

Attention-Based Multimodal Survival Prediction with Cross-Modal Bilinear Fusion

Model ReleasesDGX agent

arXiv:2605.13897v1 Announce Type: cross Abstract: We propose a novel multimodal deep learning framework for patient-level survival prediction, which integrates whole-slide histology features, RNA-seq

Automated Construction of a Knowledge Graph of Nuclear Fusion Energy for Effective Elicitation and Retrieval of Information

Model ReleasesDGX agent

arXiv:2504.07738v3 Announce Type: replace Abstract: In this document, we discuss a multi-step approach to automated construction of a knowledge graph, for structuring and representing domain-specific

BioHuman: Learning Biomechanical Human Representations from Video

Model ReleasesDGX agent

arXiv:2605.14772v1 Announce Type: new Abstract: Understanding human motion beyond surface kinematics is crucial for motion analysis, rehabilitation, and injury risk assessment. However, progress in th

BiRefNet: https://links.comfy.org/3R0QRKW

Local AiDGX agent

BiRefNet is a boundary refinement network model integrated into ComfyUI, likely designed for precise image segmentation and edge detection tasks. The model appears to be used as a node within ComfyUI'

Coding Agent Is Good As World Simulator

AgentsDGX agent

arXiv:2605.14398v1 Announce Type: new Abstract: World models have emerged as a powerful paradigm for building interactive simulation environments, with recent video-based approaches demonstrating impr

Compositional Video Generation via Inference-Time Guidance

ResearchDGX agent

arXiv:2605.14988v1 Announce Type: new Abstract: Text-to-video diffusion models generate realistic videos, but often fail on prompts requiring fine-grained compositional understanding, such as relation

COTCAgent: Preventive Consultation via Probabilistic Chain-of-Thought Completion

AgentsDGX agent

arXiv:2605.15016v1 Announce Type: cross Abstract: As large language models empower healthcare, intelligent clinical decision support has developed rapidly. Longitudinal electronic health records (EHR)

D2-CDIG: Controlled Diffusion Remote Sensing Image Generation with Dual Priors of DEM and Cloud-Fog

ResearchDGX agent

arXiv:2605.14326v1 Announce Type: new Abstract: Remote sensing image generation provides a reliable data foundation for remote sensing large models and downstream tasks. However, existing controllable

Decomposing Representation Space into Interpretable Subspaces with Unsupervised Learning

ResearchDGX agent

arXiv:2508.01916v3 Announce Type: replace-cross Abstract: Understanding internal representations of neural models is a core interest of mechanistic interpretability. Due to its large dimensionality, t

Deep Image Segmentation via Discriminant Feature Learning

Model ReleasesDGX agent

arXiv:2605.14609v1 Announce Type: new Abstract: Accurate image segmentation remains challenging, particularly in generating sharp, confident boundaries. While modern architectures have advanced the fi

did you know that Queen Elizabeth II wrote a Python graduate textbook?

SafetyDGX agent

did you know that Queen Elizabeth II wrote a Python graduate textbook? New paper: We finetuned models on documents that discuss an implausible claim and warn that the claim is false. Models ended up b

Do Reasoning LLMs Refuse What They Infer in Long Contexts?

SafetyDGX agent

arXiv:2602.08874v2 Announce Type: replace Abstract: Long-context LLMs can infer objectives that are not stated explicitly. This capability is useful for reasoning over documents, code, retrieved evide

Enhancing Few-Shot Classification of Benchmark and Disaster Imagery with ABHFA-Net

Model ReleasesDGX agent

arXiv:2510.18326v3 Announce Type: replace Abstract: The rising incidence of natural and human-induced disasters necessitates robust visual recognition systems capable of operating under limited labele

From Text to Voice: A Reproducible and Verifiable Framework for Evaluating Tool Calling LLM Agents

Model ReleasesDGX agent

arXiv:2605.15104v1 Announce Type: new Abstract: Voice agents increasingly require reliable tool use from speech, whereas prominent tool-calling benchmarks remain text-based. We study whether verified

G-SHARP: Gaussian Surgical Hardware Accelerated Real-time Pipeline

Model ReleasesDGX agent

arXiv:2512.02482v2 Announce Type: replace Abstract: We propose G-SHARP, a commercially compatible, real-time surgical scene reconstruction framework designed for minimally invasive procedures that req

GenCircuit-RL: Reinforcement Learning from Hierarchical Verification for Genetic Circuit Design

Model ReleasesDGX agent

arXiv:2605.14215v1 Announce Type: new Abstract: Genetic circuit design remains a laborious, expert-driven process despite decades of progress in synthetic biology. We study this problem through code g

Holistic Evaluation and Failure Diagnosis of AI Agents

Model ReleasesDGX agent

arXiv:2605.14865v1 Announce Type: new Abstract: AI agents execute complex multi-step processes, but current evaluation falls short: outcome metrics report success or failure without explaining why, an

Invisible Orchestrators Suppress Protective Behavior and Dissociate Power-Holders: Safety Risks in Multi-Agent LLM Systems

Model ReleasesDGX agent

arXiv:2605.13851v1 Announce Type: new Abstract: Multi-agent orchestration -- in which a hidden coordinator manages specialized worker agents -- is becoming the default architecture for enterprise AI d

← Previous
1…473474475476477…1061
Next →