AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,483
  • Agents7,562
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,175
  • Local Ai4,939
  • Model Releases23,953
  • Research20,128
  • Safety13,373
  • Syntheses17
  • Tools1,677
  • Tutorials3,405

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,483
  • Agents7,562
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,175
  • Local Ai4,939
  • Model Releases23,953
  • Research20,128
  • Safety13,373
  • Syntheses17
  • Tools1,677
  • Tutorials3,405

Source
HumanDGX agent

88,483Total entries
1Added by human
88,482Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,694 results
4 Aug 2026

Belief-Space Perception Routing under Coupled Sensor Faults and Compute Contention

Model ReleasesDGX agent

arXiv:2608.00322v1 Announce Type: cross Abstract: A robot that has to see and react on a fixed clock runs into two problems at once. Its cameras degrade in rain, mud, fog, and darkness. And the single

Beyond Edge Maps: Wavelet-Domain Conditioning for Multi-Adapter Map-to-Satellite Diffusion

Model ReleasesDGX agent

arXiv:2608.00083v1 Announce Type: new Abstract: Commercial mapping partnerships are often unavailable in low-resource regions, leaving satellite basemaps stale and motivating synthesis of satellite im

Beyond Gene Reconstruction: Learning Cell Representations through Complementary Transcriptomic Views

TutorialsDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2608.00985v1 Announce Type: new Abstract: The rapid growth of single-cell transcriptomic data has enabled the development of foundation models pretrained primarily by reconstructing masked expre

CAPEval: A Decoupled Caption Evaluation across Understanding and Generation

Model ReleasesDGX agent

arXiv:2608.02589v1 Announce Type: new Abstract: Captions serve as a primary supervision signal for both multimodal understanding and text-to-image generation. However, previous evaluations treat the c

Cardiovascular Digital Twins from Physics Based to Data Driven Approaches

ResearchDGX agent

arXiv:2608.02135v1 Announce Type: cross Abstract: Cardiovascular digital twins aim to create patient-specific computational models that evolve with clinical data to support diagnosis, prognosis, and t

CompanionBench: A Theory-Anchored, Real-World-Grounded Benchmark for AI Emotional Companionship

Model ReleasesDGX agent

arXiv:2608.02046v1 Announce Type: new Abstract: LLM companions are deployed at scale in personally consequential settings, yet poorly evaluated. Existing benchmarks use hand-authored scenarios and pro

CoNav-UAV: Cooperative Dual-Altitude Aerial Navigation via Stackelberg Learning

Model ReleasesDGX agent

arXiv:2608.01802v1 Announce Type: cross Abstract: Target-oriented vision-and-language navigation (VLN) on aerial platforms is attracting growing attention for missions such as disaster rescue, infrast

Deep Learning for Retinal Degeneration Assessment: A Comprehensive Analysis of the MARIO Challenge

Model ReleasesDGX agent

arXiv:2506.02976v4 Announce Type: replace Abstract: The MARIO challenge, held at MICCAI 2024, focused on advancing the automated detection and monitoring of age-related macular degeneration (AMD) thro

DeepConvContext: A Multi-Scale Approach to Timeseries Classification in Human Activity Recognition

ResearchDGX agent

arXiv:2505.20894v2 Announce Type: replace Abstract: Despite recognized limitations in modeling long-range temporal dependencies, Human Activity Recognition (HAR) has traditionally relied on a sliding

DeepSeek V4 Flash 0731 (Q4) now reaches 1,328 tok/s prefill and ~29 tok/s decode on one RTX PRO 6000

Model ReleasesDGX agent

I've been working on speeding up DeepSeek-V4-Flash-0731 in Krasis and have now got the long-prompt prefill quite a bit faster on a single RTX PRO 6000 96GB. These are timing-disabled internal Krasis r

Deepseek V4 flash 0731 ranks #21 on Agent Arena

Model ReleasesDGX agent

https://preview.redd.it/522fsdwvtdhh1.png?width=1200&format=png&auto=webp&s=6a6cf7a467514167a8193029dbd20fb3a9ba4f6c It ranks lower than both Sonnet 4.6 and Luna. I'd wager Luna costs in the same ball

Dense Language Generation Made Simple: Deterministic, Randomized, and Multi-Order Algorithms

TutorialsDGX agent

arXiv:2608.01320v1 Announce Type: cross Abstract: Language generation in the limit is a theoretical framework for studying how a generator can learn to produce new valid strings from a stream of posit

Detail Continuation over a Trustworthy Coarse Scale for Autoregressive Super-Resolution

Model ReleasesDGX agent

arXiv:2608.01823v1 Announce Type: new Abstract: Hallucination remains a persistent challenge in generative super-resolution (GSR), where reconstructed results may contain visually plausible yet weakly

DeVIT: Low-Power Vision Transformer Acceleration Using Delta Computation

ResearchDGX agent

arXiv:2608.01343v1 Announce Type: new Abstract: The emergence of transformer-based deep learning models has brought unprecedented performance across various domains, particularly in natural language p

Distill Where You Fail: Recovering Learning Signals of Negative RL-Groups from Adaptive Teacher Guidance

SafetyDGX agent

arXiv:2608.00782v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) has become a standard paradigm for post-training large language models (LLMs). While Group Relativ

Does Explainability Transfer? A Controlled Benchmark of Attribution Methods on Vision Transformers and CNNs

Model ReleasesDGX agent

arXiv:2608.02396v1 Announce Type: new Abstract: Most evidence on the effectiveness of explainable artificial intelligence (XAI) attribution methods has been established on convolutional neural network

Don't Offer What Can't Be Done: Deterministic Executability Gating for LLM Skill Selection at Scale

ApplicationsDGX agent

arXiv:2608.01050v1 Announce Type: cross Abstract: Production LLM agents that select from large skill libraries face a limitation that semantic relevance alone cannot resolve: a skill may match a user'

DynamicManip: Enabling Dynamic Manipulation from a Single Static Demonstration

Model ReleasesDGX agent

arXiv:2608.01452v1 Announce Type: cross Abstract: Dynamic manipulation is a critical capability for robots operating in complex and dynamic environments, where robots must interact with objects that a

Emergence Invariance: From Symbolized Thought to Interface Refinement

Model ReleasesDGX agent

arXiv:2608.01548v1 Announce Type: cross Abstract: Language can be viewed as a formalized subset of thought: a consequence-governed symbolic structure projected from wider situated cognition. Large lan

EmoScene: A Dual-space Dataset for Controllable Affective Image Generation

ResearchDGX agent

arXiv:2604.00933v2 Announce Type: replace Abstract: Text-to-image diffusion models achieve high visual fidelity, yet fine-grained affective control remains difficult because textual emotion cues often

Event ActivityNet: A Large-Scale Simulated-Event Benchmark for Untrimmed Action Understanding

Model ReleasesDGX agent

arXiv:2608.01948v1 Announce Type: new Abstract: Long-horizon event-based action understanding remains underexplored because existing datasets largely comprise short, trimmed clips, while collecting na

From Global to Local: A Scalable Benchmark for Local Posterior Sampling

Model ReleasesDGX agent

arXiv:2507.21449v2 Announce Type: replace-cross Abstract: Degeneracy is an inherent feature of the loss landscape of neural networks, but it is not well understood how stochastic gradient MCMC (SGMCMC

From We to Me: Theory Informed Narrative Shift with Abductive Reasoning

Model ReleasesDGX agent

arXiv:2603.03320v2 Announce Type: replace Abstract: Effective communication often relies on aligning a message with an audience's narrative and worldview. Narrative shift involves transforming text to

FusionRS: A Large-Scale RGB-Infrared-Style Remote Sensing Dataset for Cross-Modal Vision-Language Learning

SafetyDGX agent

arXiv:2606.17020v2 Announce Type: replace Abstract: Remote sensing vision-language models have advanced Earth observation, but available large-scale vision-language resources remain RGB-centered, leav

Future Mode Part 2: The foundation for securing agentic browsing

Model ReleasesDGX agent

Editor's Note: Our Future Mode series will give businesses insight into how Chrome Enterprise is approaching AI in the browser. Stay tuned for more blogs in this series.Future Mode Part 2: The foundat

Gecko: Fast Private Inference via Secure Public Encoder Offloading

TutorialsDGX agent

arXiv:2608.02378v1 Announce Type: new Abstract: Private inference protects both user inputs and server models during neural network inference, but existing solutions remain too slow for practical depl

Generated Images Are Easier to Forget: A Machine Unlearning Perspective for Synthetic Image Detection

ResearchDGX agent

arXiv:2608.00716v1 Announce Type: new Abstract: Robust detection of generated images is critical to counter the misuse of generative models. Existing methods primarily depend on learning from human-an

GeoArbiter: Verifiability-Guided Grounding for Remote-Sensing Multimodal LLMs

SafetyDGX agent

arXiv:2608.00877v1 Announce Type: new Abstract: Remote-sensing multimodal large language models (MLLMs) often assert facts that imagery cannot establish, such as a facility's identity or function. Coo

Global Optimization and Inference-Time Region Grafting for Agentic Workflows

Model ReleasesDGX agent

arXiv:2608.02353v1 Announce Type: new Abstract: Recent advances in agentic workflow optimization automate workflow design through task-specific workflow search or input-conditioned architecture select

HorusEye: Language as Dynamic Attention for Emergency Visual Analysis

Model ReleasesDGX agent

arXiv:2606.14741v2 Announce Type: replace Abstract: We introduce HorusEye, Language as Dynamic Attention for Emergency Visual Analysis. Our investigation followed five stages. The first one is benchma

HP-JEPA: Hierarchical Partitioning for Multi-Resolution Graph Joint-Embedding Predictive Learning

Model ReleasesDGX agent

arXiv:2608.00491v1 Announce Type: new Abstract: Graph self-supervised learning aims to learn transferable representations from large-scale unlabeled graph data. Joint-embedding predictive architecture

inclusionAI/Ling-3.0-flash weights are up on Hugging Face — MIT, BF16 plus an official FP8

Model ReleasesDGX agent

Went public in the last few minutes, both repos ungated. Ling-3.0-flash, BF16, 24 shards, ~255GB Ling-3.0-flash-fp8, official FP8, ~128GB 127.5B total, they quote 5.1B active. What jumped out at me in

Interpretability-Guided Soft Pruning of Attention Heads in Vision Transformers

TutorialsDGX agent

arXiv:2608.00264v1 Announce Type: new Abstract: Vision foundation models, such as DINOv2, learn highly expressive representations but rely on massive, opaque architectures that demand substantial comp

Learning Compositional Meta-Routing for Agentic Workflows: An Executable Benchmark

Model ReleasesDGX agent

arXiv:2608.00106v1 Announce Type: new Abstract: Agentic systems must decide not only what answer to produce, but which reasoning and execution operations should precede it. A controller may answer dir

Live and ready to build. Thanks for having us! @OpenRouter. Open weights dropping soon.⚡️

Model ReleasesDGX agent

Live and ready to build. Thanks for having us! @OpenRouter. Open weights dropping soon.⚡️ Qwen3.8 Max by @Alibaba_Qwen is live on OpenRouter. The new flagship has 2.4T parameters (95B active) and is b

LiveMem: Maintaining Memory State Continuity in Long-Running LLM Inference

Model ReleasesDGX agent

arXiv:2608.02515v1 Announce Type: new Abstract: Long-running assistants and agents consume interaction streams that eventually outgrow the context. Existing context retention, summarization, and retri

Logit-Origin Centering for Singleton Test-Time Adaptation

ApplicationsDGX agent

arXiv:2608.01074v1 Announce Type: cross Abstract: Tabular data is used extensively in many real-world use cases. Deep learning models have been developed to deal with tabular data, but generally perfo

Loop-Mamba: A Loop Mamba with Degradation-Aware and Shared Memory for Old Photo Restoration

Model ReleasesDGX agent

arXiv:2608.02346v1 Announce Type: new Abstract: Old photographs often suffer from multiple coupled degradations, including scratches, cracks, fading, blur, noise, and missing regions, severely degradi

Minute-Scale Training for Microrobot Navigation

Model ReleasesDGX agent

arXiv:2608.00854v1 Announce Type: new Abstract: Microrobots hold significant potential for various applications, where targeted navigation is a basic requirement. Deep reinforcement learning (DRL) has

Native Multilingual Chain-of-Thought Reasoning in Low-Resource Southeast Asian Languages

SafetyDGX agent

arXiv:2608.00533v1 Announce Type: new Abstract: Large Language Models have achieved substantial progress in reasoning capabilities. Yet in low-resource native settings, many suffer from cross-lingual

One Query, Many Scales: Sparse Mixture-of-Experts for Efficient Hierarchical Cross-View Geo-Localization

Model ReleasesDGX agent

arXiv:2608.01060v1 Announce Type: new Abstract: Cross-view geo-localization (CVGL) retrieves geo-tagged satellite imagery for a ground-view query. Most systems exhaustively search a flat, fixed-resolu

OTAP: Structure-Aware Optimal Transport for Evaluating Planning and Execution in Agent Trajectories

AgentsDGX agent

arXiv:2607.17082v2 Announce Type: replace-cross Abstract: Large language model agents solve tasks by generating trajectories that interleave planning, tool calls, and intermediate results. Current eva

Pc build limitations

Model ReleasesDGX agent

Here's the build I managed to scrape together System Specifications: CPU: Intel Core i7-7700K Motherboard: ASUS ROG Strix Z270-E Gaming RAM: 32GB Corsair Vengeance DDR4-3000 Storage: 1TB Crucial P5 Pl

PRISM: Privileged Probabilistic Latent Supervision for End-to-End Autonomous Driving Motion Planning

AgentsDGX agent

arXiv:2608.01201v1 Announce Type: cross Abstract: End-to-end autonomous driving (E2E AD) systems integrate perception, prediction, and planning into a single differentiable architecture. While these m

Private Generative Bootstrap via Blocking

Model ReleasesDGX agent

arXiv:2608.02480v1 Announce Type: cross Abstract: With AI systems gaining more access to individuals' information, it is important to protect privacy when reporting statistical answers. Equally import

PromptPath: Prompt-Adaptive Computational Pathways for In-Context Learning

ResearchDGX agent

arXiv:2608.02129v1 Announce Type: new Abstract: In-context learning (ICL) has attracted increasing attention for enabling models to perform new tasks using only a few ``input--output'' prompt examples

Proxy Avatar Meets Low-Rank Caching: Real-Time One-Shot Emotion-Controllable Portrait Animation

ResearchDGX agent

arXiv:2608.01978v1 Announce Type: new Abstract: Audio-driven portrait animation has advanced rapidly with diffusion-based generative models, yet real-time one-shot generation with expressive emotion c

Quick on the Uptake: Eliciting Implicit Intents from Human Demonstrations for Personalized Mobile-Use Agents

SafetyDGX agent

arXiv:2508.08645v3 Announce Type: replace Abstract: As multimodal large language models advance rapidly, the automation of mobile tasks has become increasingly feasible through the use of mobile-use a

Qwen-CUA: Native Computer Use for (almost) Everything

Model ReleasesDGX agent

arXiv:2608.02352v1 Announce Type: cross Abstract: Native computer use offers a general interface for agents to operate almost any software available to people, but requires long-horizon state tracking

RADAR: Rubric-Aware Dependency and Redundancy Analysis for LLM-as-Judge Evaluation

Model ReleasesDGX agent

arXiv:2608.01810v1 Announce Type: new Abstract: Rubric-based LLM-as-judge pipelines often assume that evaluation criteria provide independent signals. In practice, however, criteria can be behaviorall

RadPRISM: Schema-stratified radiology-report supervision for concept-disentangled image representations and visual grounding

SafetyDGX agent

arXiv:2608.00147v1 Announce Type: new Abstract: Vision-language pretraining learns rich medical image representations from radiology reports, but previous model variants commonly operate within a sing

Relax Within, Balance Across: Geometry-Guided Load Balancing for Vision-Language Mixture-of-Experts

Model ReleasesDGX agent

arXiv:2608.00574v1 Announce Type: new Abstract: Vision-language MoE batches contain different numbers of image and text tokens. Image resolution, image count, tiling, and prompt length all change this

Remember-R1: Mitigating Long-Context Visual Forgetting through Reinforcement Learning

ResearchDGX agent

arXiv:2608.01314v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) increasingly rely on long chain-of-thought reasoning for complex tasks. However, as reasoning sequences lengthe

Response Magnitude as a Dominant Signal for Held-Out CRISPRi Perturbation Effect Prediction

Model ReleasesDGX agent

arXiv:2608.00152v1 Announce Type: new Abstract: Predicting the magnitude of a CRISPRi perturbation's transcriptomic effect on held-out target genes is an important open problem in single-cell biology.

RestoreKV: Recovering Full-Cache Behavior Under Aggressive Query-Agnostic KV Cache Eviction

Model ReleasesDGX agent

arXiv:2608.01247v1 Announce Type: new Abstract: Query-agnostic KV cache eviction compresses a context once and reuses the resulting cache for arbitrary future queries, but performance can collapse und

Revisiting Generalization Across Difficulty Levels: It's Not So Easy

ResearchDGX agent

arXiv:2511.21692v2 Announce Type: replace Abstract: We investigate how well large language models (LLMs) generalize across different task difficulties, a key question for effective data curation and e

RING: Retrieval-Internalized Generation for Continual Large-Scale Knowledge Injection

Model ReleasesDGX agent

arXiv:2608.01630v1 Announce Type: new Abstract: Retrieval-augmented generation (RAG) improves factuality but adds latency and engineering overhead at serving time. We propose RING (Retrieval-Internali

Same violence, different answer: how AI responds to coercive control against women across languages

ResearchDGX agent

arXiv:2608.01436v1 Announce Type: cross Abstract: Women experiencing coercive control, a form of intimate partner violence increasingly conducted through digital devices, are turning to conversational

ScrambleToolBench: Agents Search Exhaustively Even When Their Own Map Points to the Next Step

Model ReleasesDGX agent

arXiv:2608.02358v1 Announce Type: new Abstract: To operate robustly in open-world environments, autonomous agents should be able to infer the behavior of unfamiliar systems through interaction alone,

Simulation-Based Plate-Reverb Parameter Estimation from a Single Impulse Response

Model ReleasesDGX agent

arXiv:2608.00656v1 Announce Type: cross Abstract: We present a simulation-trained, non-iterative estimator for Task A of the 1st DAFx Parameter Estimation Challenge. Each unnormalized plate-reverb imp

← Previous
1…510511512513514…1062
Next →