AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries89,023
  • Agents7,608
  • Applications5,444
  • Concepts5
  • Hardware1,854
  • Industry6,179
  • Local Ai4,968
  • Model Releases24,142
  • Research20,258
  • Safety13,456
  • Syntheses17
  • Tools1,677
  • Tutorials3,415

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries89,023
  • Agents7,608
  • Applications5,444
  • Concepts5
  • Hardware1,854
  • Industry6,179
  • Local Ai4,968
  • Model Releases24,142
  • Research20,258
  • Safety13,456
  • Syntheses17
  • Tools1,677
  • Tutorials3,415

Source
HumanDGX agent

89,023Total entries
1Added by human
89,022Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
64,155 results
10 Jul 2026

Benchmark Evaluation of Feredated Learning on Multi-organ Images

Model ReleasesDGX agent

arXiv:2607.08219v1 Announce Type: new Abstract: The privacy requirements of medical data and its substantial variations across organs and modalities hinder the clinical implementation of medical AI. F

Beyond Backpropagation: Monte Carlo Method Can Train Deep Neural Networks

Model ReleasesDGX agent

arXiv:2607.08406v1 Announce Type: new Abstract: Backpropagation (BP) dominates deep learning training, but its reliance on gradients brings inherent troubles -- vanishing and exploding gradients. The

Concept-as-Tree: A Controllable Synthetic Data Framework Makes Stronger Personalized VLMs

ResearchDGX agent

arXiv:2503.12999v4 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) have demonstrated exceptional performance in various multi-modal tasks. Recently, there has been an increasing i

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Context Graphs for Proactive Enterprise Agents

Model ReleasesDGX agent

arXiv:2607.07721v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) and agentic frameworks have advanced enterprise AI considerably, yet agents remain fundamentally reactive: they wai

DeepSWE: Measuring Frontier Coding Agents on Original, Long-Horizon Engineering Tasks

Model ReleasesDGX agent

arXiv:2607.07946v1 Announce Type: cross Abstract: DeepSWE is a benchmark of 113 original, long-horizon software engineering tasks for evaluating coding agents. Most public agentic coding benchmarks fo

Diagnosing Shape-Prior Shortcuts in Long-Range Single-Shot Fringe Projection Profilometry

Model ReleasesDGX agent

arXiv:2606.17093v2 Announce Type: replace Abstract: Learning-based single-shot fringe projection profilometry (FPP) has been studied almost entirely at close range, and the networks used are evaluated

DominoTree: Conditional Tree-Structured Drafting with Domino for Speculative Decoding

Model ReleasesDGX agent

arXiv:2607.08642v1 Announce Type: new Abstract: Speculative decoding accelerates LLM inference by drafting several tokens and verifying them in parallel. Block-diffusion drafters such as DFlash produc

Dual-Correlation Hypergraph Network for Unaligned RGBT Video Object Detection and A Large-scale Benchmark

Model ReleasesDGX agent

arXiv:2607.08191v1 Announce Type: new Abstract: RGB-Thermal (RGBT) Video Object Detection (VOD) has gained significant traction due to its ability to overcome the limitations of conventional RGB-based

Dual-Difficulty Curriculum Learning for Direct Preference Optimization

SafetyDGX agent

arXiv:2504.07856v4 Announce Type: replace Abstract: Curriculum learning enhances Direct Preference Optimization (DPO) for aligning Large Language Models (LLMs), yet existing methods rely on a one-dime

Ensemble Diversity Optimization for Subjective Supervision

SafetyDGX agent

arXiv:2607.08493v1 Announce Type: cross Abstract: Subjective NLP tasks often exhibit systematic annotator disagreement, requiring models that represent uncertainty rather than collapse it. We introduc

FedOPAL: One-Shot Federated Learning via Analytic Visual Prompt Tuning

ResearchDGX agent

arXiv:2607.08368v1 Announce Type: new Abstract: With the widespread deployment of basic models in edge intelligence, communication bandwidth has become a core bottleneck restricting the scalability of

Generalization Theory for Through-the-Wall Radar Human Activity Recognition

Model ReleasesDGX agent

arXiv:2607.08144v1 Announce Type: cross Abstract: Through-the-wall radar (TWR) human activity recognition (HAR) is important for non-line-of-sight indoor sensing, security monitoring, and emergency re

GERD: Geometric event response data generation

ApplicationsDGX agent

arXiv:2412.03259v3 Announce Type: replace Abstract: Event-based vision sensors offer high temporal resolution, high dynamic range, and low power consumption, yet event-based vision models lag behind c

GPT-5.6 is a major step forward for health intelligence. Across the lineup, we’re delivering stronger performance at lower cost: GPT-5.6 Lun…

Model ReleasesDGX agent

GPT-5.6 is a major step forward for health intelligence. Across the lineup, we’re delivering stronger performance at lower cost: GPT-5.6 Luna outperforms GPT-5.5 at its highest reasoning setting while

Graph-Regularized Deep Learning for EEG-Based Emotion Recognition with Psychologically-Grounded Label Structure

Model ReleasesDGX agent

arXiv:2607.07773v1 Announce Type: cross Abstract: EEG-based emotion recognition is critical for mental health monitoring and affective brain-computer interfaces, yet existing deep learning approaches

Hugging Face Gemma Challenge results are in! 📈 Over 6 days, more than 100 AI agents and humans collaborated to make Gemma 4 inference 5x fa…

Model ReleasesDGX agent

Hugging Face Gemma Challenge results are in! 📈 Over 6 days, more than 100 AI agents and humans collaborated to make Gemma 4 inference 5x faster on a single NVIDIA A10G GPU. - Fastest result: 491.8 TPS

ICDAR 2026 HIPE-OCRepair Competition on LLM-Assisted OCR Post-Correction for Historical Documents

Model ReleasesDGX agent

arXiv:2607.08143v1 Announce Type: cross Abstract: We present the results of HIPE-OCRepair-2026, an ICDAR competition on LLM-assisted OCR post-correction of historical documents. OCR post-correction re

It Takes a MAESTRO To Prune Bad Experts

SafetyDGX agent

arXiv:2607.08601v1 Announce Type: new Abstract: Sparsely-activated Mixture-of-Experts (MoE) language models achieve remarkable inference efficiency by activating only a small fraction of parameters pe

Jet-Long: Efficient Long-Context Extension with Dynamic Bifocal RoPE

Model ReleasesDGX agent

arXiv:2607.07740v1 Announce Type: cross Abstract: Modern LLMs are increasingly deployed in long-context applications such as retrieval-augmented generation, repository-level coding, and agentic workfl

Joint Discrete-Continuous Flow Matching for Open-Vocabulary Inverse Design of Multilayer Optical Coatings

Model ReleasesDGX agent

arXiv:2607.08392v1 Announce Type: cross Abstract: Amortized neural inverse design typically remains closed-world: component choices are fixed vocabulary tokens, coordinate grids are frozen at training

LEEVLA: Seeing What Matters in Latent Environment Evolution for Vision-Language-Action

AgentsDGX agent

arXiv:2607.08182v1 Announce Type: cross Abstract: Vision-language-action (VLA) models aim to map multimodal inputs to robot actions. However, most existing approaches struggle to cover complex dynamic

Linear Attention Architectures: Mechanisms, Trade-offs, and Cross-Layer Routing

Model ReleasesDGX agent

arXiv:2607.07953v1 Announce Type: cross Abstract: Self-attention lets each token retrieve information from the full context, but its quadratic cost in sequence length limits training and inference at

Nigeria Machinery: A Low-Resource Industrial Dataset with a Domain-Grounded Reasoning Layer

ApplicationsDGX agent

arXiv:2607.07883v1 Announce Type: new Abstract: There is relatively little, public, and model-ready data on industrial machinery for African economies. This makes it hard to do quantitative analysis o

OmniOPSD: Rationale-Privileged On-Policy Self-Distillation for Affective Computing

Local AiDGX agent

arXiv:2606.15920v2 Announce Type: replace Abstract: Reinforcement learning for multimodal large language models (MLLMs) is often hindered by severe reward sparsity in complex reasoning tasks. This cha

Open-Vocabulary Object-Goal Navigation by Generalizing Semantic Mapping with Dense CLIP

ResearchDGX agent

arXiv:2407.09016v2 Announce Type: replace Abstract: Object-oriented embodied navigation tasks require agents to locate specific objects, either defined by category or images, in unseen environments. W

OPSD-V: On-Policy Self-Distillation for Post-Training Few-Step Autoregressive Video Generators

SafetyDGX agent

arXiv:2607.08766v1 Announce Type: new Abstract: We propose OPSD-V, an on-policy self-distillation paradigm for post-training few-step autoregressive (AR) video diffusion models. Existing few-step AR v

path_boost: A Python Package for Interpretable Graph-Level Prediction using Path-Based Gradient Boosting

Model ReleasesDGX agent

arXiv:2607.07935v1 Announce Type: cross Abstract: We present path_boost, a Python package for interpretable supervised learning on graph-structured input data. The package implements PathBoost, a grad

PhyMAGIC: Physical Motion-Aware Generative Inference with Confidence-guided LLM

SafetyDGX agent

arXiv:2505.16456v3 Announce Type: replace Abstract: Recent advances in 3D content generation have amplified demand for dynamic models that are both visually realistic and physically consistent. Howeve

Pose-to-Biomechanics: Bridging 3D Human Pose Estimation and Biomechanical Attribute Prediction

Model ReleasesDGX agent

arXiv:2607.08725v1 Announce Type: cross Abstract: Recent progress in 3D human pose estimation has made markerless recovery of skeletal motion increasingly accurate and scalable. However, most pose est

Post-Training in End-to-End Autonomous Driving

SafetyDGX agent

arXiv:2607.08072v1 Announce Type: new Abstract: End-to-end models that map multimodal inputs directly to future trajectories/maneuvers have emerged as an increasingly prominent research paradigm in au

RhyMix: A Lightweight Adaptive Multi-Rhythm Network for Long-Term Time Series Forecasting

Local AiDGX agent

arXiv:2607.08234v1 Announce Type: cross Abstract: Real-world time series exhibit complex dynamics characterized by multiple simultaneous temporal patterns: short-term fluctuations, periodic seasonal c

Switch-Reasoner: Learn When to Think in Multitask Mixtures via Reinforcement Learning

TutorialsDGX agent

arXiv:2607.08572v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) often follow a fixed Think-then-Answer paradigm, which is inefficient in heterogeneous multitask settings becau

TFP: Temporally Conditioned Memory-Fusion Policies for Visuomotor Learning

Model ReleasesDGX agent

arXiv:2607.08283v1 Announce Type: new Abstract: Vision--Language--Action (VLA) policies such as pi_{0.5} and OpenVLA perform well on many manipulation tasks, but they are often reactive: the next acti

The Download: Claude’s inner workings and OpenAI’s “super app”

Model ReleasesDGX agent

This is today’s edition of The Download, our weekday newsletter that provides a daily dose of what’s going on in the world of technology. Anthropic found a hidden space where Claude puzzles over conce

The same way, we're probably one of the few AI startups with user network effects, we might become the first one with agent network effects!

Model ReleasesDGX agent

The same way, we're probably one of the few AI startups with user network effects, we might become the first one with agent network effects! Hugging Face Gemma Challenge results are in! 📈 Over 6 days,

Truncated Step-Level Sampling with Process Rewards for Retrieval-Augmented Reasoning

ResearchDGX agent

arXiv:2602.23440v4 Announce Type: replace Abstract: Reinforcement learning has emerged as an effective paradigm for training large language models to interleave reasoning with search engine calls. How

TTHE: Test-Time Harness Evolution

AgentsDGX agent

arXiv:2607.08124v1 Announce Type: cross Abstract: The behavior of an LLM agent is determined not only by the underlying model, but also by its harness: the executable program that constructs context,

UAV-OVVIS: Unmanned Aerial Vehicles Also Need Open-Vocabulary Video Instance Segmentation

Model ReleasesDGX agent

arXiv:2607.08075v1 Announce Type: new Abstract: Unmanned Aerial Vehicle (UAV) videos are widely used in traffic monitoring, urban management, and emergency rescue. However, existing UAV video percepti

Vision-Language Memory for Spatial Reasoning

ResearchDGX agent

arXiv:2511.20644v2 Announce Type: replace Abstract: Spatial reasoning is a critical capability for intelligent robots, yet current vision-language models (VLMs) still fall short of human-level perform

Workflow as Knowledge: Semantic Persistence for LLM-Mediated Workflows

SafetyDGX agent

arXiv:2607.08740v1 Announce Type: new Abstract: Large language model (LLM) applications increasingly use explicit workflows for tool use, retrieval, branching, checkpointing, and human approval. Exist

9 Jul 2026

A knowledge-augmented dataset of high-risk driving scenarios with LLM annotations for autonomous driving

Model ReleasesDGX agent

arXiv:2607.07103v1 Announce Type: new Abstract: Safe autonomous driving requires both rapid responses to common high-risk events and deeper reasoning over rare, extreme long-tail scenarios in traffic

An Adaptive Differentially Private Federated Learning Framework

TutorialsDGX agent

arXiv:2602.06838v3 Announce Type: replace Abstract: Federated learning enables collaborative model training across distributed clients while preserving data privacy. However, in practical deployments,

AnchorPrune: Relevance-Anchored Contextual Expansion for Visual Token Pruning

Local AiDGX agent

arXiv:2607.07033v1 Announce Type: cross Abstract: Large vision-language models incur substantial inference costs because high-resolution inputs introduce thousands of visual tokens, many of which are

Anomaly detection in time-series via inductive biases in the latent space of conditional normalizing flows

ApplicationsDGX agent

arXiv:2603.11756v2 Announce Type: replace Abstract: Deep generative models for anomaly detection in multivariate time-series are typically trained by maximizing observed data likelihood. However, like

ASFR-Net: Adversarial Alignment and Spatio-Frequency Refinement Network for Heterogeneous Remote Sensing Image Change Detection

Model ReleasesDGX agent

arXiv:2607.07161v1 Announce Type: new Abstract: The core challenge of heterogeneous change detection in remote sensing imagery lies in effectively decoupling genuine land-cover changes from significan

Audio Sentiment Analysis via Distillation and Cross-Modal Integration of Generated Multilingual Transcripts

ResearchDGX agent

arXiv:2607.06611v1 Announce Type: cross Abstract: Automatically recognizing the sentiment, positive or negative, from speech is a challenging task, requiring both the analysis of vocal inflections and

Breaking Database Lock-in: Agentic Regeneration of High Performance Storage Readers for Database Bypass

Model ReleasesDGX agent

arXiv:2607.07696v1 Announce Type: cross Abstract: Analytical workloads operating on data stored in external database systems face a fundamental bottleneck: data access is guarded entirely by the datab

BUS: Brain-Inspired Unsupervised Self-Reflection for Advanced Multimodal Reasoning

ResearchDGX agent

arXiv:2607.07361v1 Announce Type: new Abstract: Current Vision-Language Models (VLMs) often struggle to handle complex visual tasks that require consistent and fine-grained reasoning. Recent methods a

Cool that Grok 4.5 is #1 in some respects, even with respect to Fable 5

Model ReleasesDGX agent

Cool that Grok 4.5 is #1 in some respects, even with respect to Fable 5 SpaceXAI's Grok 4.5 takes the #1 spot on AutomationBench-AA with a score of 51%, ahead of Claude Fable 5 (49%) and Claude Opus 4

DiaLLM: An Investigation into the Robustness-Generation Gap in English Dialect Adaptation

SafetyDGX agent

arXiv:2607.07669v1 Announce Type: cross Abstract: Large language models increasingly understand dialectal English, yet still produce only standard, US-leaning English, leaving dialectal generation, th

Fixed-Gaussian Spectral Algorithms: Minimax Optimal Rates for Misspecified Learning and Transfer

Model ReleasesDGX agent

arXiv:2501.10870v2 Announce Type: replace-cross Abstract: The principal objective of this work is twofold within nonparametric regression settings: (1) to establish the minimax optimal convergence rat

FMMVCC: Fuzzy Mamba-based Multi-View Contrastive Clustering for Univariate Time Series

Model ReleasesDGX agent

arXiv:2607.07258v1 Announce Type: cross Abstract: In many realistic scenarios, large volumes of time series data are generated with limited or expensive annotations. This limitation makes supervised l

FPTQuant: Function-Preserving Transforms for LLM Quantization

ResearchDGX agent

arXiv:2506.04985v2 Announce Type: replace Abstract: Large language models (LLMs) require substantial compute, and thus energy, at inference time. While quantizing weights and activations is effective

From Beats to Breaches:How Offensive AI Infers Sensitive User Information from Playlists

Model ReleasesDGX agent

arXiv:2605.04724v2 Announce Type: replace-cross Abstract: The pervasive integration of AI has enabled Offensive AI: the exploitation of AI for malicious ends across the cyber-kill chain. A critical ma

From Text to Parameters: Predicting Item Parameters from Embedding Regularization with Reliability and Design Ceilings

Model ReleasesDGX agent

arXiv:2607.07141v1 Announce Type: new Abstract: Newly developed items must ordinarily be field tested before their psychometric properties are known, creating a cold start problem for item calibration

GPT-5.6 is now available in Devin! On FontierCode 1.1 Extended, the GPT-5.6 family stands out for pairing strong scores with excellent cost …

Model ReleasesDGX agent

GPT-5.6 is now available in Devin! On FontierCode 1.1 Extended, the GPT-5.6 family stands out for pairing strong scores with excellent cost efficiency. GPT 5.6 Sol reaches top performance at nearly ha

GPT-Live is now fully rolled out to all ChatGPT users on Go, Plus, and Pro plans. Free user rollout is in progress. Update to the latest ver…

Model ReleasesDGX agent

GPT-Live is now fully rolled out to all ChatGPT users on Go, Plus, and Pro plans. Free user rollout is in progress. Update to the latest version of the ChatGPT app on iOS or Android to try it out. Int

grok 4.5 made me give grok build a serious run today here's my honest first impression (non affiliated neutral view point): 1. grok build is…

Model ReleasesDGX agent

grok 4.5 made me give grok build a serious run today here's my honest first impression (non affiliated neutral view point): 1. grok build is a very good harness firstmate stretches harness capabilitie

Grok x Cursor

Model ReleasesDGX agent

Grok, Elon Musk's AI assistant developed by xAI, has integrated with Cursor, a popular AI-powered code editor, enabling developers to leverage Grok's capabilities for coding tasks and assistance. This

Health System Scale Semantic Search Across Unstructured Clinical Notes

Model ReleasesDGX agent

arXiv:2604.25605v2 Announce Type: replace-cross Abstract: Introduction: Semantic search, which retrieves documents based on conceptual similarity rather than keywords, offers advantages for retrieval

← Previous
1…525526527528529…1070
Next →