AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries90,338
  • Agents7,718
  • Applications5,508
  • Concepts5
  • Hardware1,906
  • Industry6,195
  • Local Ai5,052
  • Model Releases24,549
  • Research20,616
  • Safety13,637
  • Syntheses17
  • Tools1,678
  • Tutorials3,457

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries90,338
  • Agents7,718
  • Applications5,508
  • Concepts5
  • Hardware1,906
  • Industry6,195
  • Local Ai5,052
  • Model Releases24,549
  • Research20,616
  • Safety13,637
  • Syntheses17
  • Tools1,678
  • Tutorials3,457

Source
HumanDGX agent

Content type
AllBlog
90,338Total entries
1Added by human
90,337Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,187 results
Model Releases

SER: Learning to Ground Video Reasoning with Semantic Evidence Rewards

DGX agent

arXiv:2606.24726v1 Announce Type: new Abstract: Video MLLMs often struggle with fine-grained spatio-temporal reasoning, sometimes generating correct answers based on irrelevant frames or objects. Alth

model-releasesarxiv-cs-cv
24 Jun 2026
Model Releases

simonw/browser-compat-db

X Post
Paper
YouTube
Reddit
GitHub
Clear filters
DGX agent

simonw/browser-compat-db Inspired by Mozilla's new MDN MCP service - source code here - I decided to try converting their comprehensive mdn/browser-compat-data repository full of browser compatibility

model-releasessimon-willison
24 Jun 2026
Safety

Societal Alignment Frameworks Can Improve LLM Alignment

DGX agent

arXiv:2503.00069v2 Announce Type: replace-cross Abstract: Recent progress in large language models (LLMs) has focused on producing responses that meet human expectations and align with shared values -

safetyarxiv-cs-ai
24 Jun 2026
Research

Stochastic Expectation Maximization for Robust State-Space Radio Interferometric Imaging

DGX agent

arXiv:2606.23944v1 Announce Type: cross Abstract: State--space models provide a flexible framework for analyzing dynamical systems, yet they often rely on Gaussian assumptions that fail to capture hea

researcharxiv-cs-lg
24 Jun 2026
Model Releases

SURGELLM: Rethinking Multi-Task Evaluation through Task-Aware Feature Gating with Class-Balanced Normalization

DGX agent

arXiv:2606.24259v1 Announce Type: cross Abstract: Fine-tuned encoders deployed across heterogeneous NLP tasks face three compounding problems: mismatched inductive biases, class-imbalance corruption o

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

Synergizing Physically Constrained MCMC and Chemical-Informed Gaussian Processes for Reaction Network Discovery

DGX agent

arXiv:2606.23757v1 Announce Type: cross Abstract: Extracting interpretable governing equations from sparse, noisy chemical time-series data remains difficult because discrete reaction topology and con

model-releasesarxiv-cs-ai
24 Jun 2026
Agents

The Latent Bridge: A Continuous Slow-Fast Channel for Real-Time Game Agents

DGX agent

arXiv:2606.24470v1 Announce Type: new Abstract: A real-time agent for general computer use - with games as the most demanding case - must act within tens of milliseconds while still planning over seco

agentsarxiv-cs-ai
24 Jun 2026
Model Releases

TIGER: Taming Identity, Geometry, and Generative Priors for High-Quality Face Video Restoration

DGX agent

arXiv:2606.24336v1 Announce Type: new Abstract: Face Video Restoration (FVR) aims to recover high-fidelity facial videos from degraded input while preserving identity and semantic consistency across f

model-releasesarxiv-cs-cv
24 Jun 2026
Safety

Token-to-Token Alignment of Text Embeddings for Semantic Blending

DGX agent

arXiv:2606.24021v1 Announce Type: new Abstract: In modern generative models, images are specified and controlled through text prompts. In practice, images are generated from sequences of tokens derive

safetyarxiv-cs-cv
24 Jun 2026
Research

Towards Version-aware Operations and Transaction Memories for Multi-layer MeMo

DGX agent

arXiv:2606.24040v1 Announce Type: cross Abstract: MeMo proposes language models with explicit multi-layer correlation matrix memories (CMMs), where memorization, retrieval, and forgetting are architec

researcharxiv-cs-ai
24 Jun 2026
Model Releases

Try Kimi K2.7 and GLM 5.2 for free in Devin Desktop and CLI

DGX agent

Try Kimi K2.7 and GLM 5.2 for free in Devin Desktop and CLI Kimi K2.7 Code and GLM 5.2 are available in Devin Desktop and CLI Both perform strongly on FrontierCode Extended, our benchmark for real-wor

model-releasescognition-ai--x
24 Jun 2026
Model Releases

Tuning without Peeking: Provable Generalization Bounds and Robust LLM Post-Training

DGX agent

arXiv:2507.01752v4 Announce Type: replace-cross Abstract: Gradient-based optimization is the workhorse of deep learning, offering efficient and scalable training via backpropagation. However, exposing

model-releasesarxiv-cs-ai
24 Jun 2026
Local Ai

VistaRef: Boosting Visual Spatial Orientation Awareness for Pointing-to-Object Detection

DGX agent

arXiv:2606.24498v1 Announce Type: new Abstract: Grounding deictic gestures in natural images is fundamental to AR and human-robot collaboration, providing a basis for seamless spatial interaction. Whi

local-aiarxiv-cs-cv
24 Jun 2026
Research

A Generalized Formalism of Auto-Regressive Decoding for Speech Processing

DGX agent

arXiv:2606.20714v1 Announce Type: cross Abstract: In speech processing, most state-of-the-art sequence prediction models rely on auto-regressive (AR) strategies to generate output sequences based on t

researcharxiv-cs-lg
23 Jun 2026
Model Releases

A Hybrid, Multi-Layered Pipeline for Phishing and Threat Classification: Independently Validated URL and NLP Engines with a Calibrated Multi-Channel Fusion Stage

DGX agent

arXiv:2606.21690v1 Announce Type: cross Abstract: Phishing is a multi-modal threat. We present a hybrid pipeline that scores each modality with its own engine and fuses the results. Three engines are

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

A Latent Representation Learning Framework for Hyperspectral Image Emulation in Remote Sensing

DGX agent

arXiv:2603.21911v2 Announce Type: replace Abstract: Synthetic hyperspectral image (HSI) generation is essential for large-scale simulation, algorithm development, and mission design, yet traditional r

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

A Smart Classroom Behavior Analysis Framework with a New Highly Congested Classroom Dataset

DGX agent

arXiv:2606.21568v1 Announce Type: new Abstract: Student behavior detection is important for intelligent classroom analysis but remains challenging in large-class scenarios due to dense instance co-occ

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

A Standard Processing Pipeline for High-accuracy Measurement of Few-shot Regression on Laser Induced Breakdown Spectroscopy

DGX agent

arXiv:2606.21960v1 Announce Type: new Abstract: Laser-induced breakdown spectroscopy (LIBS) faces challenges in high-accuracy quantitative measurement under few-shot scenarios due to spectral noise an

model-releasesarxiv-cs-lg
23 Jun 2026
Safety

A Stitch in Time Saves Nine: Preserving Policy Compatibility Under Perception Updates in End-to-End Autonomous Driving

DGX agent

arXiv:2606.21509v1 Announce Type: new Abstract: End-to-end autonomous driving systems tightly couple perception and decision-making through latent representations. Consequently, updates to perception

safetyarxiv-cs-ro
23 Jun 2026
Hardware

An Analysis of Untrained Deep Reservoir Networks for Audio Surveillance

DGX agent

arXiv:2606.22218v1 Announce Type: cross Abstract: In this paper, we investigate untrained recurrent models from the Reservoir Computing (RC) paradigm for audio surveillance, focusing on bidirectional

hardwarearxiv-cs-lg
23 Jun 2026
Applications

An Efficient and Effective Architecture for Large-Scale Traffic Prediction via Geometry-Adaptive Square Partitioning

DGX agent

arXiv:2606.21072v1 Announce Type: new Abstract: Traffic prediction is a core task in intelligent transportation systems and urban-scale decision making. Despite the effectiveness of mainstream neural-

applicationsarxiv-cs-lg
23 Jun 2026
Model Releases

Anticipating the Optimism Gap: Predicting Distribution-Shift Degradation of RF-Impairment Detectors from In-Distribution Statistics

DGX agent

arXiv:2606.22054v1 Announce Type: cross Abstract: Detectors for GNSS radio-frequency impairments (jamming, spoofing, multipath) are usually reported with a single AUC measured on the distribution they

model-releasesarxiv-cs-lg
23 Jun 2026
Research

Boundary-by-Mask: Few-Shot Instance Segmentation with Mask-Conditioned Boundary Learning for Texture-Poor Industrial Parts

DGX agent

arXiv:2606.21594v1 Announce Type: new Abstract: Recent advances in large pre-trained models have led to remarkable progress in instance segmentation on general images. However, industrial scenarios re

researcharxiv-cs-cv
23 Jun 2026
Model Releases

BranchShine: Compact Raw-Audio-to-IPA Transcription with a RoPE E-Branchformer Encoder

DGX agent

arXiv:2606.22824v1 Announce Type: new Abstract: Speech-to-IPA transcription is useful when the desired output is pronunciation rather than orthographic text, but competitive multilingual systems are o

model-releasesarxiv-cs-lg
23 Jun 2026
Local Ai

BYOK is now live in the GitHub Copilot App! Works with @ollama, foundry, and any OAI completions or Anthropic compatible messages endpoint. …

DGX agent

GitHub Copilot App now supports Bring Your Own Key (BYOK) functionality, allowing users to integrate local and third-party AI models including Ollama, Foundry, and any OpenAI-compatible or Anthropic-c

local-aiollama--x
23 Jun 2026
Research

Category-Adaptive Cross-Modal Semantic Refinement and Transfer for Open-Vocabulary Multi-Label Recognition

DGX agent

arXiv:2412.06190v2 Announce Type: replace Abstract: Benefiting from the generalization capability of CLIP, recent vision language pre-training (VLP) models have demonstrated the ability to capture a w

researcharxiv-cs-cv
23 Jun 2026
Safety

CFPO: Counterfactual Policy Optimization for Multimodal Reasoning

DGX agent

arXiv:2606.23206v1 Announce Type: new Abstract: Large Vision-Language Models (LVLMs) have demonstrated remarkable capabilities in multimodal reasoning. However, prevailing reinforcement learning (RL)

safetyarxiv-cs-cv
23 Jun 2026
Research

Cloak: Zero-Shot Cross-Embodiment Manipulation by Masking the End-Effector from the VLA

DGX agent

arXiv:2606.22836v1 Announce Type: new Abstract: We present Cloak, a training recipe that endows a Vision-Language-Action (VLA) model with zero-shot cross-embodiment transfer by cloaking the end-effect

researcharxiv-cs-ro
23 Jun 2026
Model Releases

CodePercept: Code-Grounded Visual STEM Perception for MLLMs

DGX agent

arXiv:2603.10757v2 Announce Type: replace Abstract: When MLLMs fail at Science, Technology, Engineering, and Mathematics (STEM) visual reasoning, a fundamental question arises: is it due to perceptual

model-releasesarxiv-cs-cv
23 Jun 2026
Local Ai

CoDMD: Copula-aware Distribution Matching Distillation for Fast Video Generation

DGX agent

arXiv:2606.21982v1 Announce Type: new Abstract: Few-step distillation for video diffusion models has attracted significant attention, driven by the urgent demand for efficient deployment in real-world

local-aiarxiv-cs-cv
23 Jun 2026
Research

Coherence Under Commitment: Probing Generalization and Vacuous Memorization in LLM Logical Reasoning

DGX agent

arXiv:2606.21083v1 Announce Type: cross Abstract: Large language models (LLMs) deployed for logical reasoning in knowledge-intensive domains exhibit a subtle but critical failure: coherence can be vac

researcharxiv-cs-lg
23 Jun 2026
Research

Data Selection Through Iterative Self-Filtering for Vision-Language Settings

DGX agent

arXiv:2606.23611v1 Announce Type: new Abstract: The availability of large amounts of clean data is paramount to training neural networks. However, at large scales, manual oversight is impractical, res

researcharxiv-cs-cv
23 Jun 2026
Safety

Distribution-Aware Diffusion-LLM for Robust Ultra-Long-Term Time Series Forecasting

DGX agent

arXiv:2606.23391v1 Announce Type: new Abstract: Time series forecasting is a fundamental machine learning task. Recent work has explored Large Language Models (LLMs) for this purpose due to their stro

safetyarxiv-cs-lg
23 Jun 2026
Model Releases

Do Location Encoders Capture Spatial Effects? A GeoShapley Benchmark Across Scales

DGX agent

arXiv:2606.23453v1 Announce Type: new Abstract: Location encoders transform geographic coordinates into high dimensional embeddings for downstream machine learning, but it is unclear how well these re

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

DR-Mamba: Automatic Inference-Time Domain Adaptation for Document Image Binarization via Sample-Conditioned Detail-Background Suppression

DGX agent

arXiv:2606.22625v1 Announce Type: new Abstract: Degraded document image binarization is sensitive to domain shifts caused by paper aging, bleed-through, stains, shadows, and uneven illumination, and t

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

DrivingVoxels: Compositional Sparse Voxel Rasterization for Dynamic Driving Scene Reconstruction

DGX agent

arXiv:2606.23031v1 Announce Type: new Abstract: Reconstructing dynamic urban scenes remains challenging due to the unbounded nature of driving environments and the presence of multiple dynamic objects

model-releasesarxiv-cs-cv
23 Jun 2026
Local Ai

Enabling Cloud-Level Accuracy in Edge AI through IoT Data Preprocessing

DGX agent

arXiv:2606.22496v1 Announce Type: cross Abstract: Large language models (LLMs) offer a natural-language interface for interpreting Internet of Things (IoT) sensor data in smart environments; however,

local-aiarxiv-cs-lg
23 Jun 2026
Agents

Enhancing Creativity in 3D Generative Design via a TRIZ-Inspired Text-to-CAD Framework

DGX agent

arXiv:2606.21378v1 Announce Type: new Abstract: Recent advances in large language models (LLMs) have demonstrated significant potential in supporting engineering design tasks, including computer-aided

agentsarxiv-cs-lg
23 Jun 2026
Model Releases

ENVS: Environment-Native Verified Search for Long-Horizon GUI Agents

DGX agent

arXiv:2606.22948v1 Announce Type: cross Abstract: As multimodal agents move from interface understanding to real software control, successful trajectory discovery in live desktop environments becomes

model-releasesarxiv-cs-cv
23 Jun 2026
Tools

Experimenting with the proposed Cross-Origin Storage API in Transformers.js

DGX agent

This article explores the Cross-Origin Storage API and its implementation within Transformers.js, a JavaScript library for machine learning models. It likely discusses how this API enables secure cros

toolshugging-face
23 Jun 2026
Model Releases

Explore-Execute Chain: Towards an Efficient Structured Reasoning Paradigm

DGX agent

arXiv:2509.23946v3 Announce Type: replace Abstract: Many LLMs plan before they act, yet planning and execution are often still entangled in one long generation trace, enforced only through prompts, or

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Factor-Aware Mixture-of-Experts with Pretrained Encoder for Combinatorial Generalization

DGX agent

arXiv:2606.21100v1 Announce Type: new Abstract: The integration of pretrained encoders with diffusion policies has become a dominant paradigm for visual robotic manipulation. However, it still struggl

model-releasesarxiv-cs-ro
23 Jun 2026
Model Releases

Factored Gossip DiLoCo: Reducing Blocking Communication in DiLoCo

DGX agent

arXiv:2606.22768v1 Announce Type: new Abstract: To make large-scale distributed training practical outside high-bandwidth datacenters, we must reduce blocking, high-volume synchronization. While DiLoC

model-releasesarxiv-cs-lg
23 Jun 2026
Research

Faithful Grounded Visual Reasoning via Learned Proxy-Tokens

DGX agent

arXiv:2606.23354v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have achieved remarkable success in Visual Question Answering (VQA), yet their 'black-box' nature hinders deplo

researcharxiv-cs-cv
23 Jun 2026
Model Releases

FirstPass: Grounding AI Scientific Judgment in Multi-Round Editorial Outcomes

DGX agent

arXiv:2606.20769v1 Announce Type: cross Abstract: AI systems for peer review fail on three fronts: they train on Computer Science and Machine Learning venues alone, ignore the iterative dialogue that

model-releasesarxiv-cs-lg
23 Jun 2026
Applications

FLFL: Federated Latent Factor Learning for Private Recovery of Spatio-Temporal Signals

DGX agent

arXiv:2606.23091v1 Announce Type: new Abstract: Wireless sensor network (WSNs) stands out as a burgeoning and promising domain in intelligent sensing. Owing to various factors such as sudden sensor ma

applicationsarxiv-cs-lg
23 Jun 2026
Safety

FOCA: Future-Oriented Conditioning for Data-Efficient Vision-Language-Action Adaptation

DGX agent

arXiv:2606.20867v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models enable general-purpose robotic control via large-scale multimodal pretraining, yet their effectiveness under few-sho

safetyarxiv-cs-cv
23 Jun 2026
Research

From Markov to Laplace: How Mamba In-Context Learns Markov Chains

DGX agent

arXiv:2502.10178v2 Announce Type: replace Abstract: While transformer-based language models have driven the AI revolution thus far, their computational complexity has spurred growing interest in viabl

researcharxiv-cs-lg
23 Jun 2026
← Previous
1…683684685686687…1359
Next →