AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
AllBlog
91,060Total entries
1Added by human
91,059Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,794 results
Safety

VIDS-Seg: Towards Reliable Uncertainty Quantification in Pediatric Cardiac Ultrasound Segmentation

DGX agent

arXiv:2608.10903v1 Announce Type: new Abstract: Reliable clinical deployment of machine learning requires models that know when they are likely to fail, particularly for subgroups underrepresented in

safetyarxiv-cs-cv
12 Aug 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

VoxSumm: A Multilingual Corpus of Long-Form Spoken News for Joint Summarization and Translation

DGX agent

arXiv:2608.10359v1 Announce Type: cross Abstract: As information increasingly traverses linguistic boundaries, users require concise cross-lingual representations of long-form content. Nevertheless, l

model-releasesarxiv-cs-cl
12 Aug 2026
Safety

Your LLM, Your Style: Behavioral Mode Axes for LLM Behavioral Control

DGX agent

arXiv:2608.10703v1 Announce Type: cross Abstract: Large language models (LLMs) increasingly act in interactive settings where their behavioral styles affect user experience, safety, and downstream dec

safetyarxiv-cs-ai
12 Aug 2026
Model Releases

1 Day in and I feel okay saying Muse-Glimmer-30B finally beats 3.6-27B for the size in some use-cases

DGX agent

A few things right off the bat: it reasons very efficiently. Like Grok 4.5 levels of efficient thinking it quantizes very well. My first few tests with iq3_xxs were better than Qwen/Gemma behaved at t

model-releasesr-localllama
11 Aug 2026
Research

A Domain-Structured Ensemble Framework for Perioperative Outcome Prediction Using Electronic Health Record Data

DGX agent

arXiv:2608.08920v1 Announce Type: new Abstract: Perioperative risk prediction models are often limited by narrow surgical populations, incomplete intraoperative data, poor calibration, and limited int

researcharxiv-cs-lg
11 Aug 2026
Model Releases

Adversarial Attacks on Deep OCR Systems

DGX agent

arXiv:2608.07636v1 Announce Type: cross Abstract: Deep-OCR (DeepSeek-OCR) advances document recognition by treating the visual modality as an optical compression medium, enabling long-context OCR at l

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Analysis and experiments of the dissipative Twistcar: direction reversal and asymptotic approximations

DGX agent

arXiv:2506.19112v3 Announce Type: replace Abstract: Underactuated wheeled vehicles are commonly studied as nonholonomic systems with periodic actuation. Twistcar is a classical example inspired by a r

model-releasesarxiv-cs-ro
11 Aug 2026
Model Releases

Autorubric: A Unifying Framework for Rubric-Based LLM Evaluation on Non-Verifiable Tasks

DGX agent

arXiv:2603.00077v3 Announce Type: replace-cross Abstract: Rubric-based LLM judges have become indispensable for evaluating and optimizing systems on non-verifiable tasks, where success cannot be reduc

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Benchmarking In-context Experiential Learning Through Repeated Product Recommendations

DGX agent

arXiv:2511.22130v2 Announce Type: replace Abstract: To navigate ever-shifting real-world environments, agents must grapple with incomplete knowledge and adapt their strategies through experience. Howe

model-releasesarxiv-cs-lg
11 Aug 2026
Model Releases

Beyond Direct Identifiers: Probabilistic Privacy Risk Estimation for Privacy-Conscious LLM Query Delegation

DGX agent

arXiv:2608.09140v1 Announce Type: cross Abstract: Recent work on protecting privacy during user-LLM interactions often focuses on direct, explicit identifiers: the personally-identifiable information

model-releasesarxiv-cs-cl
11 Aug 2026
Local Ai

Beyond the Plane: Coupling Planar Vehicle Dynamics with Three-Dimensional Road Geometry

DGX agent

arXiv:2608.09402v1 Announce Type: new Abstract: Simulation is crucial for developing and testing autonomous driving systems. In particular, the development of localization and control algorithms relie

local-aiarxiv-cs-ro
11 Aug 2026
Model Releases

Claude's watermark probably doesn't work how you think. As the CTO of GPTZero, I'll explain how Anthropic, Google and OpenAI are building te…

DGX agent

Claude's watermark probably doesn't work how you think. As the CTO of GPTZero, I'll explain how Anthropic, Google and OpenAI are building text watermarking in this brief explainer and whether it can b

model-releasesemad-mostaque--x
11 Aug 2026
Model Releases

Describe-to-Score: A text-guided framework for image complexity assessment

DGX agent

arXiv:2509.16609v2 Announce Type: replace Abstract: Accurately assessing image complexity (IC) is essential for many vision tasks, yet existing approaches rely almost exclusively on visual features an

model-releasesarxiv-cs-cv
11 Aug 2026
Model Releases

Detecting Clear Contact Lenses for Iris Recognition: A Two-Stage Mask-Guided Attention Approach

DGX agent

arXiv:2608.08977v1 Announce Type: cross Abstract: This work focuses on the impact and detection of clear contact lenses in the context of iris recognition. While the detection of cosmetic or patterned

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Domain-Aware Pruning: Sparsity and Domain Generalization via Regularized Probabilistic Masking

DGX agent

arXiv:2608.08624v1 Announce Type: new Abstract: Domain generalization (DG) and neural network pruning are conventionally treated as distinct objectives, targeting out-of-distribution (OOD) robustness

model-releasesarxiv-cs-lg
11 Aug 2026
Safety

Dynamic Distribution-Aware Uncertainty Tracking in Vision-Language Representation Learning

DGX agent

arXiv:2608.09011v1 Announce Type: new Abstract: Uncertainty Quantification (UQ) aims to measure the reliability of model predictions, serving as a critical safeguard for deploying Vision-Language Mode

safetyarxiv-cs-lg
11 Aug 2026
Research

EEG Foundation Challenge: From Cross-Task to Cross-Subject EEG Decoding

DGX agent

arXiv:2506.19141v3 Announce Type: replace-cross Abstract: Current electroencephalogram (EEG) decoding models are typically trained on small numbers of subjects performing a single task. Here, we intro

researcharxiv-cs-lg
11 Aug 2026
Model Releases

Effect of Abstractions and Prompting Strategies on LLM-Guided High-Performance Optimizations

DGX agent

arXiv:2608.08085v1 Announce Type: cross Abstract: Code performance optimization is a vital aspect of modern software development, as it enables faster response times and reduced resource usage. These

model-releasesarxiv-cs-ai
11 Aug 2026
Research

Enhancing Knowledge Tracing through Leakage-Free and Recency-Aware Embeddings

DGX agent

arXiv:2508.17092v2 Announce Type: replace-cross Abstract: Knowledge Tracing (KT) aims to predict a student's future performance based on their sequence of interactions with learning content. Many KT m

researcharxiv-cs-ai
11 Aug 2026
Research

Estimating Uncertainty in Galaxy Morphology Classification

DGX agent

arXiv:2608.08398v1 Announce Type: new Abstract: Astronomers classify galaxy morphology to investigate cosmic evolution. While deep foundation models are increasingly utilized in Galaxy Morphology Clas

researcharxiv-cs-ai
11 Aug 2026
Local Ai

Evidence-Calibrated Runtime Reconstruction for Agent Skills Across Heterogeneous Coding Agents

DGX agent

arXiv:2608.08793v1 Announce Type: new Abstract: Agent Skills package reusable instructions and assets for tool-using language-model agents. Progressive loading creates failure boundaries poorly repres

local-aiarxiv-cs-cl
11 Aug 2026
Model Releases

ExtractBench is one of the most comprehensive benchmarks for real-world document extraction. ✅ It covers 4869 pages, across 67 document type…

DGX agent

ExtractBench is one of the most comprehensive benchmarks for real-world document extraction. ✅ It covers 4869 pages, across 67 document types, spanning 8 real-world domains: finance, energy, gov, auto

model-releasesjerry-liu--x
11 Aug 2026
Model Releases

FedA2L: Adaptive layer-wise learning rate adjustment in decentralized federated learning

DGX agent

arXiv:2608.09208v1 Announce Type: cross Abstract: Decentralized intelligence systems with heterogeneous devices and limited coordination increasingly rely on decentralized federated learning (DFL). Ho

model-releasesarxiv-cs-ai
11 Aug 2026
Local Ai

FedTVD: Balancing Data Quality and Quantity for Robust Federated Learning

DGX agent

arXiv:2608.09221v1 Announce Type: cross Abstract: Federated Learning (FL) enables collaborative model training across distributed client devices while preserving data privacy. However, FL faces signif

local-aiarxiv-cs-ai
11 Aug 2026
Model Releases

From Benchmark Performance to Tool Deployment: Human-in-the-Loop Anomaly Detection

DGX agent

arXiv:2608.07770v1 Announce Type: cross Abstract: Automated anomaly detection methods often report strong performance on curated academic benchmarks, but their behavior under real-world industrial con

model-releasesarxiv-cs-cv
11 Aug 2026
Safety

From Rebound to Remedy: Understanding and Mitigating Reward Hacking via Representation Engineering

DGX agent

arXiv:2604.01476v2 Announce Type: replace-cross Abstract: Reinforcement learning for LLMs is vulnerable to reward hacking, where models exploit shortcuts to maximize reward without solving the intende

safetyarxiv-cs-cl
11 Aug 2026
Research

Full-Feature versus Limited-Input Machine Learning for Residential Energy Estimation: A Comparative Analysis of RECS and ResStock Under Realistic Input Constraints

DGX agent

arXiv:2608.09255v1 Announce Type: new Abstract: Residential energy estimates are often needed before detailed envelope characteristics, equipment efficiencies, infiltration, sensor, or billing data ar

researcharxiv-cs-lg
11 Aug 2026
Applications

How Simple Can It Get? From Interpretable Equations to Readable Rules for Financial Decision Making

DGX agent

arXiv:2608.09433v1 Announce Type: cross Abstract: In regulated domains such as finance, a model that cannot be explained cannot be deployed, yet many interpretable classifiers defeat their own purpose

applicationsarxiv-cs-ai
11 Aug 2026
Model Releases

I built a weird, low-power llama.cpp server using an Intel N100 + RTX 5060Ti

DGX agent

Everything started with the sudden death of my old ASRock J1900. While looking for the perfect ITX replacement, I stumbled upon the Chinese CW-NAS-ADLN-K motherboard, which looked perfect on paper: In

model-releasesr-localllama
11 Aug 2026
Model Releases

Instability of LLM Pre-Pretraining: It Doesn't Always Help. An Investigation on Multiple Languages

DGX agent

arXiv:2608.08800v1 Announce Type: new Abstract: Pretraining LLMs on artificial languages ('pre-pretraining') is a technique that could reportedly increase token efficiency by 33%, i.e., save up to 33%

model-releasesarxiv-cs-cl
11 Aug 2026
Model Releases

InstructionCrafter: Generating Consistent and High-Fidelity Visual Instructions

DGX agent

arXiv:2608.08460v1 Announce Type: new Abstract: Given textual task instructions, generating step-by-step visual instructions as an image sequence requires the simultaneous satisfaction of multiple pro

model-releasesarxiv-cs-cv
11 Aug 2026
Safety

Investigating Multimodal Informativity under Different Partner Visibility Conditions in Video-Mediated Dialogue

DGX agent

arXiv:2608.08915v1 Announce Type: new Abstract: Situated language use is multimodal and embodied. For example, gestures can carry information that is absent or underspecified in the speech signal, yet

safetyarxiv-cs-cl
11 Aug 2026
Safety

Is CLIP Cross-Eyed? Revealing and Mitigating Center Bias in the CLIP Family

DGX agent

arXiv:2604.05971v2 Announce Type: replace-cross Abstract: Recent research has shown that contrastive vision-language models such as CLIP often lack fine-grained understanding of visual content. While

safetyarxiv-cs-cl
11 Aug 2026
Model Releases

JUMP-lite: Compact, reproducible benchmarking of cell representations

DGX agent

arXiv:2608.07632v1 Announce Type: cross Abstract: Image-based profiling captures rich phenotypic signatures for drug discovery and functional genomics. Large public datasets like JUMP Cell Painting no

model-releasesarxiv-cs-cv
11 Aug 2026
Tutorials

Let Geometry GUIDE: Layer-wise Unrolling of Geometric Priors in Multimodal LLMs

DGX agent

arXiv:2604.05695v2 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) have achieved remarkable progress in 2D visual tasks but still struggle to understand physical space in rea

tutorialsarxiv-cs-cv
11 Aug 2026
Research

LITEWAY: LIghtweight HAR via Temporal Efficient highWAY

DGX agent

arXiv:2608.09421v1 Announce Type: cross Abstract: Wearable human activity recognition (HAR) remains challenging due to the computational and energy constraints of deep learning models on resource-limi

researcharxiv-cs-ai
11 Aug 2026
Model Releases

Llama-CPP Parallel Agents --> fine for decode, but one agent's prefill will grind all other agents to a halt

DGX agent

Testing with 3-5 agents. Decode performance is superb, however if one performs a web search and needs to process a few thousand tokens, ALL other agents will grind to a halt: I've tried tuning a littl

model-releasesr-localllama
11 Aug 2026
Model Releases

LogiShot: Logically Coherent Cross-Shot Video Generation

DGX agent

arXiv:2608.08820v1 Announce Type: new Abstract: Generating cross-shot videos that are logically connected is essential for content creation. Currently, most cross-shot video-generation workflows, such

model-releasesarxiv-cs-cv
11 Aug 2026
Model Releases

Marrying Optimal Transport and ODEs for Unified Continuous-Time 4D Reconstruction and Tracking

DGX agent

arXiv:2608.09613v1 Announce Type: new Abstract: Existing unified 4D reconstruction and point tracking approaches typically rely on heuristic interpolations or just predict at integer timestamps, lacki

model-releasesarxiv-cs-cv
11 Aug 2026
Model Releases

MasDrift: Benchmarking Authorization Preservation Across Multi-Agent Architectures

DGX agent

arXiv:2608.07556v1 Announce Type: cross Abstract: Multi-agent systems (MAS) decompose long-horizon tasks across supervisors and subagents, but delegated goals do not necessarily carry their original a

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Matrix-free Neural Preconditioner for the Dirac Operator in Lattice Gauge Theory

DGX agent

arXiv:2509.10378v2 Announce Type: replace-cross Abstract: Linear systems arise in generating samples and in calculating observables in lattice quantum chromodynamics~(QCD). Solving the Hermitian posit

model-releasesarxiv-cs-lg
11 Aug 2026
Model Releases

Muse-Glimmer 30B Hits ~280 t/s in Real Production Coding

DGX agent

These numbers were captured during a real feature implementation task in Next.js and Nest.js (adding a theme switching system across components). The structural predictability of UI/state refactoring

model-releasesr-localllama
11 Aug 2026
Model Releases

My conversation with @ericvishria of Benchmark. Eric has spent a decade investing across software and hardware, backing companies like Firew…

DGX agent

My conversation with @ericvishria of Benchmark. Eric has spent a decade investing across software and hardware, backing companies like Fireworks, Sierra, Sunday Robotics, and Cerebras. This one is abo

model-releasessonya-huang--x
11 Aug 2026
Model Releases

OmnilingualGAIA2: Evaluating the Multilingual Gap in Frontier AI Agents

DGX agent

arXiv:2608.08775v1 Announce Type: new Abstract: Agentic benchmarks aim to measure how well AI agents plan, search, execute, and recover within realistic multi-tool environments, but they are almost ex

model-releasesarxiv-cs-cl
11 Aug 2026
Model Releases

OpenVisTool: An Open Recipe for Synthesizing Instructive Visual Tool-Use Trajectories

DGX agent

arXiv:2608.08557v1 Announce Type: new Abstract: Visual tool use has emerged as a fundamental capability for multimodal agents to actively acquire evidence beyond a fixed image encoding. The prevailing

model-releasesarxiv-cs-cl
11 Aug 2026
Research

Opportunity Is Not Realizability: Selection-Valid Diagnostics for Multi-LLM Routing

DGX agent

arXiv:2608.08265v1 Announce Type: new Abstract: Oracle routing measures how much a pool of language models could gain from per-query selection, but the diagnostic has two flaws: testing against a best

researcharxiv-cs-lg
11 Aug 2026
Safety

Planning/RL for a stochastic single-player merge puzzle: afterstates, previewed chance events, and long-horizon throughput [D]

DGX agent

I am working on an AI for a small single-player merge puzzle and would appreciate pointers to related algorithms, papers, or existing implementations. It resembles 2048 in its action -> afterstate ->

safetyr-machinelearning
11 Aug 2026
Local Ai

PosBridge: Multi-View Positional Embedding Transplant for Identity-Aware Image Editing

DGX agent

arXiv:2508.17302v2 Announce Type: replace Abstract: Localized subject-driven image editing aims to seamlessly integrate user-specified objects into target scenes. As generative models continue to scal

local-aiarxiv-cs-cv
11 Aug 2026
← Previous
1…574575576577578…1371
Next →