AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,617
  • Agents7,497
  • Applications5,365
  • Concepts5
  • Hardware1,816
  • Industry6,151
  • Local Ai4,900
  • Model Releases23,593
  • Research19,967
  • Safety13,267
  • Syntheses17
  • Tools1,674
  • Tutorials3,365

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,617
  • Agents7,497
  • Applications5,365
  • Concepts5
  • Hardware1,816
  • Industry6,151
  • Local Ai4,900
  • Model Releases23,593
  • Research19,967
  • Safety13,267
  • Syntheses17
  • Tools1,674
  • Tutorials3,365

Source
HumanDGX agent

87,617Total entries
1Added by human
87,616Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,987 results
12 Aug 2026

Leveraging Human Reading Behavior for Keyphrase Extraction: A Webcam-based Eye-tracking Corpus

ResearchDGX agent

arXiv:2608.10688v1 Announce Type: new Abstract: Purpose: Keyphrases are statistically and semantically important textual units that can also attract readers' attention during comprehension. However, e

LLM Agents Factory: Retrieval of Domain-Specific LLM Agents

AgentsDGX agent

arXiv:2608.09934v1 Announce Type: cross Abstract: Large language model (LLM) agents improve task performance by decomposing problems into role-specialized behaviors. However, their practical deploymen

Local AI is exploding! Transformers.js, that we've been building @huggingface for the past three years has now become the most popular open-…

Local AiDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Local AI is exploding! Transformers.js, that we've been building @huggingface for the past three years has now become the most popular open-source library to run AI models directly in your browser and

MAP-Graph: Provenance-Aware Shared Memory for Multi-Agent Workflows

Model ReleasesDGX agent

arXiv:2608.10509v1 Announce Type: new Abstract: Shared memory helps language-model agents reuse information across long workflows, yet relevant evidence may not be admissible for a particular agent or

MMArt A Multi-Perspective Multimodal Dataset for Visual Art Understanding

ResearchDGX agent

arXiv:2608.10706v1 Announce Type: new Abstract: Recent vision-language models demonstrate impressive general visual understanding, yet their art interpretation remains shallow: they describe surface c

Multilingual Embedding Probes Fail to Generalize Across Learner Corpora

ResearchDGX agent

arXiv:2604.07095v2 Announce Type: replace Abstract: Do multilingual embedding models encode a language-general representation of proficiency? We investigate this by training linear and non-linear prob

NullEdit: Stealthy Image Protection via VLM Condition Redirection

Model ReleasesDGX agent

arXiv:2608.10870v1 Announce Type: new Abstract: Modern image editors combine vision-language models (VLMs) with diffusion transformer backbones to modify a single reference image according to instruct

Once Poisoned, Arbitrarily Controlled: A Programmable Backdoor in VLMs

TutorialsDGX agent

arXiv:2608.10959v1 Announce Type: new Abstract: Existing vision-language model (VLM) backdoors are usually treated as static vulnerabilities: one-to-one and N-to-N attacks bind one or more triggers to

RadFusion: Towards Threshold-Controllable Radiology Report Generation

ResearchDGX agent

arXiv:2608.10505v1 Announce Type: new Abstract: Automated radiology report generation is advancing rapidly in response to the shortage of radiologists, yet unlike a perception model, existing generati

ReRound: Reconstructive Rounding to Resolve Midpoint Ambiguity in Calibration-Free LLM Quantization

Model ReleasesDGX agent

arXiv:2608.11045v1 Announce Type: cross Abstract: ReRound (Reconstructive Rounding) is a post-training quantization method that addresses the midpoint ambiguity inherent in standard round-to-nearest (

SAR2Agri: Learning SAR Intensity Representations for Agricultural Monitoring

Model ReleasesDGX agent

arXiv:2608.11142v1 Announce Type: new Abstract: Agricultural monitoring faces unique challenges, arising from the landscape's complex temporal, phenological, and climate dynamics, yet monitoring them

Selective Prediction Reduces the Negative Effects of Automation Bias Overall but Increases False Negatives

SafetyDGX agent

arXiv:2508.07617v2 Announce Type: replace-cross Abstract: AI has the potential to augment human decision making. However, even high-performing models can produce inaccurate predictions when deployed.

Simplex Relaxation for Discrete Diffusion

ResearchDGX agent

arXiv:2608.10615v1 Announce Type: new Abstract: Discrete diffusion models for categorical generation are defined by a corruption kernel, which determines the intermediate state space and the associate

Speaking of fine-tuning, we’ve got support on @axolotl_ai. Start training North Micro Vision right away, no hardware required. Find their do…

Model ReleasesDGX agent

Speaking of fine-tuning, we’ve got support on @axolotl_ai. Start training North Micro Vision right away, no hardware required. Find their docs here: https://docs.axolotl.ai/docs/models/cohere-north-mi

ThinkRetrieve: Retrieval-Augmented Reasoning Traces for Test-Time Scaling

TutorialsDGX agent

arXiv:2608.10928v1 Announce Type: new Abstract: Large Reasoning Models (LRMs) improve performance by allocating additional inference-time compute to generate extended chain-of-thought reasoning. Howev

Topological Feasibility Guarantees for Differentiable Predictive Control

SafetyDGX agent

arXiv:2608.10332v1 Announce Type: cross Abstract: Differentiable predictive control (DPC), a self-supervised learning approach for approximating explicit model predictive control (MPC) policies, offer

Try Qwen-Image-3.0 on @openart_ai! 🎨👀

Model ReleasesDGX agent

Try Qwen-Image-3.0 on @openart_ai! 🎨👀 Qwen Image 3.0 is now on OpenArt ✨ The most Real Qwen image model yet. Native text across 12 languages, precise 10px type, and full interfaces like web pages, gam

Uncertainty-Aware Deep Learning for Genomics Applications: Insights from an Empirical Study

ResearchDGX agent

arXiv:2608.11054v1 Announce Type: new Abstract: Deep learning models have emerged as the standard computational tool for a wide range of applications in genomics. Yet, uncertainty quantification (UQ)

UniProbe: A Learnable Token-Level Hallucination Detector for Large VLMs using Multi-Structural Internal Representations

Local AiDGX agent

arXiv:2608.10835v1 Announce Type: new Abstract: Large Vision-Language Models (LVLMs) achieve impressive visual reasoning and dialogue capabilities, yet frequently hallucinate content unsupported by th

UPAIR: Diagnosing Reasoning States via Uncertainty-Progress Alignment for Selective Intervention

SafetyDGX agent

arXiv:2607.17188v2 Announce Type: replace Abstract: While test-time scaling improves the problem-solving ability of large reasoning models (LRMs) through additional inference-time computation, it can

UT-ACA: Uncertainty-Triggered Adaptive Context Allocation for Long-Context Inference

Model ReleasesDGX agent

arXiv:2603.18446v2 Announce Type: replace Abstract: Long-context inference remains challenging for large language models due to attention dilution and out-of-distribution degradation. Context selectio

Validated Synthetic Patient Generation for Small Longitudinal Cohorts: Coagulation Dynamics Across Pregnancy

ResearchDGX agent

arXiv:2604.07557v2 Announce Type: replace Abstract: Small longitudinal cohorts, common in maternal health, rare diseases, and early-phase trials, limit computational modeling because enrollment is slo

VidForensics-M1: Meta-Detection Reinforcement Learning with Verifiable Temporal Grounding for AI-Generated Video Forensics

Local AiDGX agent

arXiv:2608.11201v1 Announce Type: new Abstract: Recent advances in video generation models have significantly improved the realism of synthetic videos, blurring the boundary between generated and auth

We wrote a 36-page ArXiv whitepaper on ExtractBench 🧑‍🔬 , our effort to create the most comprehensive, schema-guided, real-world document …

Model ReleasesDGX agent

We wrote a 36-page ArXiv whitepaper on ExtractBench 🧑‍🔬 , our effort to create the most comprehensive, schema-guided, real-world document extraction benchmark. It’s extremely detailed and covers every

When Vision Becomes Text: Visual Token Pruning via Cross-Modal Residual Guidance in VLMs

Local AiDGX agent

arXiv:2608.10489v1 Announce Type: new Abstract: Abundant visual information strengthens vision-language model (VLM) perception, yet massive visual tokens raise inference costs. Existing visual token p

Withholding the Completing Chunk: Deterministic Pair-Completion Guardrails for Streaming LLM Output

Model ReleasesDGX agent

arXiv:2608.10279v1 Announce Type: cross Abstract: Streaming language-model output creates a release-timing problem: complete-response moderation acts after streamed text has escaped, whereas repeated

Yukon: Open Innovation for Frontier Research Open Innovation has been the driving value at Eigen Labs. But I’ve often struggled with where i…

AgentsDGX agent

Yukon: Open Innovation for Frontier Research Open Innovation has been the driving value at Eigen Labs. But I’ve often struggled with where it is actually better than a great closed team. Finally we fo

11 Aug 2026

12GB VRAM gang, what's our plan?

Model ReleasesDGX agent

Seems like we're limited to qwen finetuned MoEs for now. Looking at the current landscape - focus seems to be on dense models (muse glimmer 30b, qwen 3.8 27b) for smaller setups. Is upgrading to 24GB

A Control Function Framework for Mitigating Position Bias in Learning to Rank Systems

Model ReleasesDGX agent

arXiv:2506.06989v3 Announce Type: replace-cross Abstract: Learning-to-rank (LTR) systems commonly depend on implicit feedback, such as user clicks, because it is easy to collect and can serve as a val

A Tight Lower Bound for Smooth Nonconvex Stochastic Optimization with Bounded Gradient Noise

Model ReleasesDGX agent

arXiv:2608.09004v1 Announce Type: cross Abstract: We prove a sharp lower bound for smooth nonconvex stochastic optimization with uniformly bounded gradient noise. In the (K=1) fresh-sample model, ever

Accurate Ensembles, Fragile Narratives: Multi-Scale Stacking and a Fidelity Audit of LLM-Generated Explanations for Credit Risk

ResearchDGX agent

arXiv:2608.08126v1 Announce Type: cross Abstract: Credit scoring increasingly relies on models whose decision logic cannot be read off their parameters, in tension with supervisory expectations that a

AdaDINO: Pair-Aware In-Backbone Adaptation of Frozen DINO for Efficient Remote Sensing Change Detection

Model ReleasesDGX agent

arXiv:2608.07982v1 Announce Type: new Abstract: Vision foundation models (VFMs) such as DINO are pretrained for single-image representation, whereas remote sensing change detection requires reasoning

AeroDPO: Unleashing Lightweight UAV Navigation with High-Fidelity Perception and Automated Preference Optimization

Model ReleasesDGX agent

arXiv:2608.07557v1 Announce Type: cross Abstract: Vision-Language Navigation for Unmanned Aerial Vehicles (UAV-VLN) requires rapid and reactive control in complex 3D environments. Recent minimalist en

AeroReformer2: Spoken-Query Referring Segmentation for Aerial Images

Model ReleasesDGX agent

arXiv:2608.08874v1 Announce Type: new Abstract: Spoken language offers a natural, hands-free interface for specifying an arbitrary target in dense remote-sensing imagery, yet existing referring remote

AI Evaluation Should Measure Verification Cost, Not Correctness Alone

Model ReleasesDGX agent

arXiv:2608.08709v1 Announce Type: new Abstract: The reliability of AI generative models is typically measured by output correctness, yet in practice it depends on the effort required to verify those o

Anchor-Based AI Approach for Pre-Crash Object Detection Utilizing Micro-Doppler Signatures in Automotive Radar

Model ReleasesDGX agent

arXiv:2608.08701v1 Announce Type: new Abstract: Advanced automated driving presents significant potential to improve modern automotive safety systems, but it depends highly on the reliable activation

AQUA20: A Benchmark Dataset for Underwater Species Classification under Challenging Conditions

Model ReleasesDGX agent

arXiv:2506.17455v3 Announce Type: replace Abstract: Robust visual recognition in underwater environments remains a significant challenge due to complex distortions such as turbidity, low illumination,

Automating Deception: Scalable Multi-Turn LLM Jailbreaks

Model ReleasesDGX agent

arXiv:2511.19517v3 Announce Type: replace-cross Abstract: Multi-turn conversational attacks, which leverage psychological principles like Foot-in-the-Door (FITD), where a small initial request paves t

AutoRefine: Compiling Trajectories into Validated Typed Agent Artifacts

Model ReleasesDGX agent

arXiv:2601.22758v2 Announce Type: replace Abstract: Large language model agents repeatedly encounter related tasks, yet systems that learn from trajectories commit every lesson to one predefined artif

Avalon-ToM-Bench: Evaluating Fine-Grained Theory of Mind via Asymmetric Game Mechanics

Model ReleasesDGX agent

arXiv:2608.09638v1 Announce Type: new Abstract: Theory of Mind (ToM) is essential for agent interactions, yet existing evaluations either rely on static scenarios that oversimplify mental-state reason

Back to the Future: A workbook time machine for spread sheet creation benchmarks

Model ReleasesDGX agent

arXiv:2608.07873v1 Announce Type: new Abstract: We introduce the workbook time machine, a pipeline that automatically creates benchmarks evaluating the ability of language models to create derived obj

Beyond Naturalness: Probing Automated Text-To-Speech Evaluators on Linguistically Grounded Dimensions

Model ReleasesDGX agent

arXiv:2608.09930v1 Announce Type: cross Abstract: Automated Text-to-Speech (TTS) evaluation methods (Mean Opinion Score (MOS) predictors and Audio Large Language Models (Audio-LLM) judges) are expecte

Can Graph Learning Learn Circuits?

Model ReleasesDGX agent

arXiv:2608.08536v1 Announce Type: new Abstract: Circuit localization is a mechanistic interpretability task whose goal is to identify a sparse subgraph of a transformer's computation graph sufficient

CAP: A Scalable Benchmark for Evaluating Cross-Site Browser Agents with Complex Actions and Perception

Model ReleasesDGX agent

arXiv:2608.08392v1 Announce Type: new Abstract: Large language models are increasingly deployed as autonomous agents that interact with the web through browsers. While recent progress has been driven

Circuit Fine-Tuning for Compute-Efficient Transformer Adaptation

Model ReleasesDGX agent

arXiv:2608.08336v1 Announce Type: new Abstract: Parameter-Efficient Fine-Tuning (PEFT) has become the de facto standard for adapting Vision Transformers (ViTs) to downstream tasks. While parameter cou

Claude will apply invisible watermarks to AI text and images

Model ReleasesDGX agent

Anthropic has pledged to start marking Claude-generated text and images with machine-readable data, in an effort to comply with European rules for AI transparency. 'Generated text will carry embedded

Control-Diverse Reinforcement Fine-Tuning: Decoupling the Shared Control Bottleneck of RL Post-Training

Model ReleasesDGX agent

arXiv:2608.08224v1 Announce Type: new Abstract: Reinforcement learning post-training unlocks complex reasoning in LLMs. Yet benchmark scores reveal only whether a model improved, not what changed insi

Coupled Graph--Policy Distillation for Personalized Medication Safety in Older Adults with Multimorbidity

Model ReleasesDGX agent

arXiv:2608.09443v1 Announce Type: new Abstract: Large language model (LLM) agents can support medication review between clinical visits, but safe choices for older adults with multimorbidity depend on

CresOWLve: Benchmarking Creative Problem-Solving Over Real-World Knowledge

Model ReleasesDGX agent

arXiv:2604.03374v2 Announce Type: replace-cross Abstract: Creative problem-solving requires combining multiple cognitive abilities, including logical reasoning, lateral thinking, analogy-making, and c

Decentralized Nonconvex Composite Federated Learning with Gradient Tracking and Momentum

Model ReleasesDGX agent

arXiv:2504.12742v2 Announce Type: replace Abstract: Decentralized Federated Learning (DFL) enables collaborative model training without relying on a central server. When local objectives are nonconvex

DeepSeek V4 Flash 0731 is now available to fine-tune on Together AI. Specialize it for coding, tool use, and your own domain with SFT or DPO…

Model ReleasesDGX agent

DeepSeek V4 Flash 0731 is now available to fine-tune on Together AI. Specialize it for coding, tool use, and your own domain with SFT or DPO, then deploy the fine-tuned model on Together AI for produc

Depth-Aware Implicit Neural Representation Priors for 3D Gravity Inversion

Model ReleasesDGX agent

arXiv:2608.08959v1 Announce Type: new Abstract: Gravimetry images subsurface density contrasts associated with geological structures, geothermal systems, and intrusive bodies. Recovering a three-dimen

Does a Toehold Make a Bidder Bolder? Preemption and Multiplicity in Multi-Round Takeover Auctions

Model ReleasesDGX agent

arXiv:2608.08407v1 Announce Type: cross Abstract: A bidder can quietly buy a stake in a company before making an offer for it. That stake, a toehold, is supposed to pay for itself twice: it makes the

Efficient Human-Contact Representation for Human-Scene Interaction

Model ReleasesDGX agent

arXiv:2608.09388v1 Announce Type: new Abstract: Human-scene interaction is an active research topic with several industrial applications in virtual reality, gaming, robotics, and surveillance. Despite

EvoTrustRAG: Evolution-Aware Conflict Attribution and Evidence Handling for Reliable Retrieval-Augmented Generation

Model ReleasesDGX agent

arXiv:2608.07933v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) improves the factuality of large language models with external knowledge, yet conflicting evidence remains a fund

Expert-Guided Multimodal Fusion for Unified Emotion and Sentiment Analysis

Model ReleasesDGX agent

arXiv:2601.07565v2 Announce Type: replace-cross Abstract: Multimodal emotion understanding requires the integration of heterogeneous data sources, including text, audio, and visual modalities, while s

From Chains to DAGs: Probing the Graph Structure of Reasoning in LLMs

ResearchDGX agent

arXiv:2601.17593v3 Announce Type: replace Abstract: Recent progress in large language models has renewed interest in how multi-step reasoning is represented internally. While prior work often treats r

Governing the KV Cache: Preventing Timing Side-Channel Leakage in Multi-Tenant LLM Inference

Model ReleasesDGX agent

arXiv:2608.09225v1 Announce Type: cross Abstract: The key-value (KV) cache is the primary throughput optimization in modern large language model (LLM) inference, enabling prefix reuse across requests.

High-Quality Exposure Correction with Diffusion-Based Image Generation Priors

ResearchDGX agent

arXiv:2608.08720v1 Announce Type: new Abstract: Although most existing exposure correction methods achieve high fidelity, they often place excessive focus on overall pixel-wise accuracy, making it cha

HindsightBench: A Black-Box Behavioral Audit Protocol for Parametric Hindsight in Time-Indexed LLM Decision Tasks

ResearchDGX agent

arXiv:2607.18867v2 Announce Type: replace-cross Abstract: Large language models leak parametric knowledge of what followed a historical date into decision tasks indexed by that date -- not necessarily

← Previous
1…376377378379380…1050
Next →