AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
Human
84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
84,569 results
4 Aug 2026

Qwen-CUA: Native Computer Use for (almost) Everything

Model ReleasesDGX agent

arXiv:2608.02352v1 Announce Type: cross Abstract: Native computer use offers a general interface for agents to operate almost any software available to people, but requires long-horizon state tracking

Qwen-Image-3.0-Pro has made a massive leap from the previous generation, now ranking #5 globally. Appreciate the recognition! We will keep b…

Model ReleasesDGX agent

Qwen-Image-3.0-Pro has made a massive leap from the previous generation, now ranking #5 globally. Appreciate the recognition! We will keep building.🚀 More exciting news from @Alibaba_Qwen: Qwen-Image-

Qwen3.8-Max, better and cheaper. Try it out. 👀

Model ReleasesDGX agent

Qwen3.8-Max, better and cheaper. Try it out. 👀 We ran a test between the new Qwen3.8-Max, Opus 5 and GPT-5.6 Sol. 3 models. same prompt. one-shot with the /design command. Reviewed gameplay features,

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Qwen3.8-Max is available in Hermes Agent now! Let's build! 🚀🚀

Model ReleasesDGX agent

Qwen 3.8‑Max, Alibaba.Qwen’s latest large‑language model, has been added to Hermes Agent and can currently be accessed at a 20 % discount. The update aims to streamline integration of the model for de

QWRF-Net: A Quantum-Wavelet Framework with Rectified Flow for Short-Term Precipitation Nowcasting

ResearchDGX agent

arXiv:2608.01626v1 Announce Type: new Abstract: Short-term precipitation nowcasting is important for hydrometeorological early warning, especially when intense convective rainfall may trigger urban fl

RADAR Perception for Dynamic Obstacle Avoidance onboard small-scale Quadrotor UAVs

ResearchDGX agent

arXiv:2608.01855v1 Announce Type: new Abstract: Fast dynamic obstacle avoidance (DOA) on uncrewed aerial vehicles (UAVs) demands not only low-latency control and actuation but also reliable perception

RADAR: Rubric-Aware Dependency and Redundancy Analysis for LLM-as-Judge Evaluation

Model ReleasesDGX agent

arXiv:2608.01810v1 Announce Type: new Abstract: Rubric-based LLM-as-judge pipelines often assume that evaluation criteria provide independent signals. In practice, however, criteria can be behaviorall

RadPRISM: Schema-stratified radiology-report supervision for concept-disentangled image representations and visual grounding

SafetyDGX agent

arXiv:2608.00147v1 Announce Type: new Abstract: Vision-language pretraining learns rich medical image representations from radiology reports, but previous model variants commonly operate within a sing

RadYOLO: Computationally Efficient 3D Object Detection and Segmentation in CT and MRI

Local AiDGX agent

arXiv:2608.00508v1 Announce Type: new Abstract: Object detection and segmentation in three-dimensional medical images is a very active area of research. However, most proposed deep learning models car

RAG Strategies for Natural Language-Based SQL Query and REST API Call Generation

ApplicationsDGX agent

arXiv:2602.07086v2 Announce Type: replace-cross Abstract: Enterprise software systems commonly expose business functionality through both relational databases and REST APIs. Accessing these interfaces

RAGOCR: Optical Compression of Retrieval-Augmented Text via Visual Representation

ResearchDGX agent

arXiv:2608.00765v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) has become essential for knowledge-intensive question answering, yet scaling RAG pipelines remains challenging due

Rake-Compress Riccati Recursions for Parallel Scenario-Tree Model Predictive Control

SafetyDGX agent

arXiv:2608.01332v1 Announce Type: cross Abstract: Scenario-tree model predictive control (MPC) represents future information by a rooted tree and optimizes a nonanticipative policy over that tree. Num

RamanPFN: learning from Raman spectral structure with a tabular foundation model

Model ReleasesDGX agent

arXiv:2608.02157v1 Announce Type: new Abstract: Raman spectroscopy enables non-destructive, label-free molecular characterization across materials science, biomedicine and process monitoring. Predicti

Randomized Algorithms for Learning Partitions with Near Optimal Query Complexity in Constant Rounds

TutorialsDGX agent

arXiv:2608.02176v1 Announce Type: cross Abstract: We study the round complexity of learning a hidden partition P of an n-element universe using PAIR queries: PAIR(x,y) tells us whether x and y belong

Ranking Image Fusion the Way Humans Do: A Learned Pairwise Preference Metric for Infrared-Visible Fusion Assessment

Model ReleasesDGX agent

arXiv:2608.01301v1 Announce Type: new Abstract: Infrared-visible image fusion (IVIF) has no ideal fused reference, so fusion algorithms are routinely ranked by scalar objective metrics that formalize

RAP: KV-Cache Compression via RoPE-Aligned Pruning

Model ReleasesDGX agent

arXiv:2602.02599v4 Announce Type: replace Abstract: Long-context inference in large language models (LLMs) is bottlenecked by the memory and compute of the key-value (KV) cache. Structured pruning is

Rapid Embodiment Adaptation for Quadrupedal Locomotion

SafetyDGX agent

arXiv:2608.01506v1 Announce Type: cross Abstract: Humans readily adapt their movements as their bodies change through aging, injury, or load carrying, but learning-based robot policies often break whe

ReACT-CLIP: Response-Aware Test-Time Defense for Vision--Language Models

ResearchDGX agent

arXiv:2608.01067v1 Announce Type: new Abstract: Training-free test-time defenses offer a practical way to improve the adversarial robustness of CLIP-style vision--language models without modifying the

Real-Time Detection and Repair of LLM Agent Failures

Model ReleasesDGX agent

arXiv:2608.02464v1 Announce Type: cross Abstract: LLM agents fail mid-episode -- they loop, cascade tool errors, drift off goal, fabricate results, or silently absorb corrupted content -- and the stan

Real-Time Visual Obstruction Detection in Surgical Augmented Reality

Model ReleasesDGX agent

arXiv:2608.00232v1 Announce Type: new Abstract: Surgical augmented reality (AR) can provide contextual guidance by overlaying virtual annotations, tool cues, and procedural information onto the surgic

ReasonCast: Towards Explainable Time Series Forecasting with Reasoning

Model ReleasesDGX agent

arXiv:2608.01875v1 Announce Type: cross Abstract: Most time series (TS) models are specialized for a single task, either understanding (i.e., returning text answers about a TS) or generation (i.e., re

Reassessing the Feasibility of PPG-Based Non-Invasive Blood Glucose Level Estimation

ApplicationsDGX agent

arXiv:2608.01820v1 Announce Type: cross Abstract: Non-invasive blood glucose level (BGL) estimation from photoplethysmography (PPG) holds great promise for wearable health monitoring, but results acro

ReBRAC-v2: The Return of the King

ResearchDGX agent

arXiv:2608.01205v1 Announce Type: new Abstract: Recent offline reinforcement learning methods increasingly rely on expressive generative policies and specialized value-guidance mechanisms. We ask whet

Recompute or Reuse? Diagnosing and Mitigating Textual Shortcuts in VLM Self-Reflection

ResearchDGX agent

arXiv:2608.01930v1 Announce Type: new Abstract: Vision-language models (VLMs) are expected to revise their reasoning when visual evidence changes. Failures to do so are often attributed to insufficien

Reconstruction-Shift Discrimination via Mask-Guided Latent Diffusion for Medical Anomaly Detection

Local AiDGX agent

arXiv:2608.00444v1 Announce Type: new Abstract: Unsupervised medical anomaly detection learns normal anatomical patterns from healthy training images and identifies deviations at test time. Reconstruc

Recursive Gaussian Processes and the Bayesian Brain

ResearchDGX agent

arXiv:2608.00503v1 Announce Type: cross Abstract: Predictive coding offers a powerful framework for cortical computation, yet scalable implementations that respect both Bayesian exactness and neurobio

Recursive Vision Language Models for General Symbolic Reasoning

Model ReleasesDGX agent

arXiv:2608.01534v1 Announce Type: new Abstract: Hard symbolic-reasoning tasks such as Sudoku, maze pathfinding, and ARC remain challenging for LLMs due to their fixed-depth autoregressive reasoning, w

Red Hat leads open-source project to automate AI governance

SafetyDGX agent

IBM Corp.’s Red Hat subsidiary today announced the formation of asago, an open-source community project intended to turn artificial intelligence governance policies into operational controls that can

Refine Drugs, Don't Complete Them: Uniform-Source Discrete Flows for Fragment-Based Drug Discovery

Model ReleasesDGX agent

arXiv:2509.26405v2 Announce Type: replace Abstract: We introduce InVirtuoGen, a discrete flow generative model for fragmented SMILES for de novo and fragment-constrained generation, and target-propert

REFLEX: Rethinking MoE Inference as Refinement-Aware Compute Allocation in Diffusion Language Models

Model ReleasesDGX agent

arXiv:2608.01784v1 Announce Type: cross Abstract: Mixture-of-experts (MoE) models increase parameter capacity by activating only a small subset of experts for each token. This conditional-computation

ReFP-AD: Rectified Flow Preconditioning for Energy-Based Anomaly Detection

ResearchDGX agent

arXiv:2608.01793v1 Announce Type: new Abstract: Unified anomaly detection requires modeling highly heterogeneous normal data without access to anomalous samples. While foundation models like DINOv2 pr

Regularized Schrodinger Bridge via Distortion-Perception Perturbation for High-Fidelity Speech Enhancement

SafetyDGX agent

arXiv:2511.11686v4 Announce Type: replace Abstract: Speech enhancement (SE) requires high-fidelity reconstruction of clean speech that preserves linguistic and paralinguistic cues while maintaining hi

Relative Parameter Importance in Task-Agnostic Replay-Free Continual Learning

Model ReleasesDGX agent

arXiv:2608.00630v1 Announce Type: new Abstract: Achieving continual learning (CL) with deep neural networks requires balancing stability and plasticity while enabling knowledge transfer. In this work,

Relax Within, Balance Across: Geometry-Guided Load Balancing for Vision-Language Mixture-of-Experts

Model ReleasesDGX agent

arXiv:2608.00574v1 Announce Type: new Abstract: Vision-language MoE batches contain different numbers of image and text tokens. Image resolution, image count, tiling, and prompt length all change this

Releasing Pokee-Isaac 28B — the world’s first real 10M-token context frontier-class agentic model, deployable on a single GPU (starting from…

Model ReleasesDGX agent

Releasing Pokee-Isaac 28B — the world’s first real 10M-token context frontier-class agentic model, deployable on a single GPU (starting from RTX 4090 or equivalent). New proprietary non-decoder-only a

Remember OpenAI tripled its score on ARC-AGI with a harness fix? We fixed harness for Kimi K3 on CyberGym E2E: 2x better at vulnerability de…

Model ReleasesDGX agent

Remember OpenAI tripled its score on ARC-AGI with a harness fix? We fixed harness for Kimi K3 on CyberGym E2E: 2x better at vulnerability detection, 3x at patching K3 is SOTA cyber-defense model you c

Remember-R1: Mitigating Long-Context Visual Forgetting through Reinforcement Learning

ResearchDGX agent

arXiv:2608.01314v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) increasingly rely on long chain-of-thought reasoning for complex tasks. However, as reasoning sequences lengthe

ReMiX-MAE: Learning Missing-Channel Cross-Modal Representations from RGB-Only Clinical Facial Videos for Sympathetic-Mediated Pain Assessment

ApplicationsDGX agent

arXiv:2608.02561v1 Announce Type: new Abstract: Automated pain assessment in real clinics is limited by scarce clinically grounded facial video data with weak labels (often sequence-level self-report)

Representation Transfer of Foundation Models for Ultra-Widefield Retinal Imaging

ResearchDGX agent

arXiv:2608.00586v1 Announce Type: new Abstract: Despite the widespread adoption of foundation models as feature extractors for medical imaging, relatively little is understood about how different pret

Residual-Based Adaptive Kalman Filtering for Legged Robot State Estimation

Model ReleasesDGX agent

arXiv:2608.02316v1 Announce Type: new Abstract: State estimation is a key component in model-based control of walking robots and, more broadly, applicable wherever hidden variables must be inferred. T

Response Magnitude as a Dominant Signal for Held-Out CRISPRi Perturbation Effect Prediction

Model ReleasesDGX agent

arXiv:2608.00152v1 Announce Type: new Abstract: Predicting the magnitude of a CRISPRi perturbation's transcriptomic effect on held-out target genes is an important open problem in single-cell biology.

Restore Text First, Enhance Image Later: Two-Stage Scene Text Image Super-Resolution with Glyph Structure Guidance

TutorialsDGX agent

arXiv:2510.21590v3 Announce Type: replace Abstract: Current image super-resolution methods show strong performance on natural images but distort text, creating a fundamental trade-off between image qu

RestoreKV: Recovering Full-Cache Behavior Under Aggressive Query-Agnostic KV Cache Eviction

Model ReleasesDGX agent

arXiv:2608.01247v1 Announce Type: new Abstract: Query-agnostic KV cache eviction compresses a context once and reuses the resulting cache for arbitrary future queries, but performance can collapse und

Rethinking and formalising the state across languages: a unified computational learning theory account

ResearchDGX agent

arXiv:2608.00523v1 Announce Type: new Abstract: The linguistic notion of state has traditionally been restricted to the construct (annexation) state of Afroasiatic languages and treated as a language-

Rethinking Federated Graph Foundation Models: A Graph-Language Alignment-based Approach

Model ReleasesDGX agent

arXiv:2601.21369v2 Announce Type: replace Abstract: Recent studies of federated graph foundational models (FedGFMs) break the idealized and untenable assumption of having centralized data storage to t

Rethinking IRSTD: Single-Point Supervision Guided Encoder-only Framework is Enough for Infrared Small Target Detection

ResearchDGX agent

arXiv:2604.05363v2 Announce Type: replace Abstract: Infrared small target detection (IRSTD) aims to separate small targets from clutter backgrounds. Extensive research is dedicated to the pixel-level

Rethinking Personalized Reward Modeling for LLMs under Preference Heterogeneity via Group-Debiased Federated Learning

Local AiDGX agent

arXiv:2608.01556v1 Announce Type: new Abstract: Large language models are increasingly aligned to human preferences via reward modeling, but user preference data are sensitive and often cannot be cent

Rethinking PPG-based Sleep Staging: Datasets, Metrics, and Benchmarks

ResearchDGX agent

arXiv:2608.00943v1 Announce Type: cross Abstract: Automated sleep staging assigns discrete stage labels to successive time epochs throughout an overnight recording; conventionally each window spans at

Rethinking Pretraining for Specialized Design Data: Evidence from the JONES-19 Cultural Design Dataset

ResearchDGX agent

arXiv:2608.00135v1 Announce Type: cross Abstract: Design and architectural archives encode expert human knowledge in graphical formats, providing a critical testbed for design-inspired Machine Learnin

Rethinking Total Absorption Gamma Spectroscopy Deconvolution: Supervised Machine Learning vs Response-Matrix Methods

ResearchDGX agent

arXiv:2608.00090v1 Announce Type: cross Abstract: The extraction of eta-feeding distributions in Total Absorption gamma-ray Spectroscopy constitutes a challenging inverse problem, particularly in nucl

Rethinking Video Token Compression with a Global Codebook: Learning Once, Compressing Everywhere

ResearchDGX agent

arXiv:2608.01271v1 Announce Type: new Abstract: Video large language models (Video-LLMs) represent videos as dense sequences of visual tokens, whose length grows with the temporal and spatial extent o

ReTouch: Empowering Contact-Rich Dexterous Manipulation with Online-Refined Tactile Prediction

ApplicationsDGX agent

arXiv:2608.01824v1 Announce Type: new Abstract: Fusing tactile signals has proven effective for contact-rich manipulation, enabling robots to perceive contact states and adapt to rapidly changing phys

Retrieval Augmented Biomedical Question Answering with Weak Question Recovery and Neural Reranking for BioASQ Task 14b

ResearchDGX agent

arXiv:2608.01468v1 Announce Type: new Abstract: This work presents DS@GT ARC BioASQ team's work for a biomedical question answering pipeline, integrating multi-source query expansion, neural reranking

Retrieval-Augmented Interpretable Learning: Towards Task-Specific Zero-Shot Models in Healthcare

ApplicationsDGX agent

arXiv:2607.17508v2 Announce Type: replace Abstract: We introduce Retrieval-Augmented Interpretable Learning (RAIL), a probabilistic meta-learning framework for zero-shot generation of task-specific in

Retrieval-Based Cross-Domain Generalization in Optical Networks via Global Features

ResearchDGX agent

arXiv:2608.00044v1 Announce Type: cross Abstract: We propose a retrieval-based framework for crossdomain quality-of-transmission (QoT) estimation that leverages transferable feature representations wh

Reusing Rollouts under Policy Lag: Prefix-Normalized Policy Optimization for LLM Reinforcement Learning

Model ReleasesDGX agent

arXiv:2608.01418v1 Announce Type: cross Abstract: Autoregressive rollout generation is a major computational cost in reinforcement learning for large language models. Reusing each rollout batch for ad

Revisiting Generalization Across Difficulty Levels: It's Not So Easy

ResearchDGX agent

arXiv:2511.21692v2 Announce Type: replace Abstract: We investigate how well large language models (LLMs) generalize across different task difficulties, a key question for effective data curation and e

RF-HOI: Recognize Human-Object Interaction with Radio Frequency Signals

ApplicationsDGX agent

arXiv:2608.00289v1 Announce Type: cross Abstract: Recognizing Human-Object Interactions (HOI) is essential for intelligent systems, underpinning applications in virtual and augmented reality, embodied

RH-RAG: Trustworthy Long-Form Generation for Privacy-Constrained Settings

Local AiDGX agent

arXiv:2608.01311v1 Announce Type: new Abstract: Generating long-form content from extensive internal reports remains challenging for organizations operating under strict privacy and security constrain

RHEA: Reliability-Harmonized Reconstruction and Assignment for Robust Multimodal-Attributed Graph Clustering

ResearchDGX agent

arXiv:2608.00621v1 Announce Type: new Abstract: Multimodal-attributed graphs (MAGs), whose nodes carry heterogeneous attributes such as text and images over a relational structure, have become a funda

← Previous
1…121122123124125…1410
Next →