AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries92,413
  • Agents7,866
  • Applications5,605
  • Concepts5
  • Hardware1,963
  • Industry6,240
  • Local Ai5,175
  • Model Releases25,275
  • Research21,121
  • Safety13,951
  • Syntheses17
  • Tools1,680
  • Tutorials3,515

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries92,413
  • Agents7,866
  • Applications5,605
  • Concepts5
  • Hardware1,963
  • Industry6,240
  • Local Ai5,175
  • Model Releases25,275
  • Research21,121
  • Safety13,951
  • Syntheses17
  • Tools1,680
  • Tutorials3,515

Source
HumanDGX agent

Content type
92,413Total entries
1Added by human
92,412Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
66,933 results
Applications

EduDial: Constructing a Large-scale Multi-turn Teacher-Student Dialogue Corpus

DGX agent

arXiv:2510.12899v3 Announce Type: replace Abstract: Recently, several multi-turn dialogue benchmarks have been proposed to evaluate the conversational abilities of large language models (LLMs). As LLM

applicationsarxiv-cs-cl
27 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

Efficient Training with Foresight: Multi-Token Auxiliary Supervision for Autoregressive Image Generation

DGX agent

arXiv:2608.25386v1 Announce Type: new Abstract: Autoregressive (AR) image generation has shown strong potential for scalable high-fidelity synthesis by modeling images as discrete token sequences. How

researcharxiv-cs-cv
27 Aug 2026
Model Releases

EncoTESS: Age-Sensitive Encodings from Raw TESS Light Curves

DGX agent

arXiv:2608.25019v1 Announce Type: cross Abstract: Main sequence stars of spectral types late F through M exhibit systematic variability in photometric light curves, particularly when they are young. R

model-releasesarxiv-cs-lg
27 Aug 2026
Model Releases

GraftSR: Grafting Authentic Textures for Real-World Image Super-Resolution via Identical-Instance Guidance

DGX agent

arXiv:2608.25334v1 Announce Type: new Abstract: Diffusion-based real-world image super-resolution (SR) achieves impressive perceptual quality but inherently suffers from severe texture hallucination.

model-releasesarxiv-cs-cv
27 Aug 2026
Model Releases

If you maintain a skill library for long-horizon agents, this one is worth your time. (bookmark it) It discusses one of most common topics I…

DGX agent

If you maintain a skill library for long-horizon agents, this one is worth your time. (bookmark it) It discusses one of most common topics I get asked about these days. It shares some good ideas on ho

model-releasesdair-ai--x
27 Aug 2026
Research

InteractGesture: Progressive Chunk Guidance for Continuous Streaming Co-Speech Gesture Control

DGX agent

arXiv:2608.25734v1 Announce Type: new Abstract: Co-speech gesture generation has made significant progress toward realistic full-body motion from speaker audio, yet existing models lack fine-grained s

researcharxiv-cs-cv
27 Aug 2026
Model Releases

Key Point Analysis Needs Structure Recovery: Task Definition, Dataset Diagnosis, and a Structure-Aware Benchmark

DGX agent

arXiv:2608.25854v1 Announce Type: new Abstract: Key Point Analysis (KPA) aims to identify a concise set of key points that summarize a collection of arguments together with their prevalence. We argue

model-releasesarxiv-cs-cl
27 Aug 2026
Model Releases

Krea2 Turbo Distill 4 step LoRA (NOT a New Checkpoint... YET to share / still mid training) ... but fun 2-step extreme experiment with surprising results... (OUT OF TRAINING SPEC, which is 4 steps!)

DGX agent

I am sure for those of you who have been following my 4 step Krea 2 Turbo LoRA, you would know from my previous posts the work of progress I have been sharing with you ( if not see here - https://www.

model-releasesr-stablediffusion
27 Aug 2026
Model Releases

LibriBrain100: One Hundred Hours of Broad and Deep MEG Data for Neural Speech Decoding at Scale

DGX agent

arXiv:2608.25204v1 Announce Type: cross Abstract: We introduce LibriBrain100, a large-scale MEG dataset for speech decoding designed from the ground up for reproducible, standardised evaluation. Libri

model-releasesarxiv-cs-cl
27 Aug 2026
Local Ai

LLMTrace: A Corpus for Classification and Fine-Grained Localization of AI-Written Text

DGX agent

arXiv:2509.21269v2 Announce Type: replace Abstract: The widespread use of human-like text from Large Language Models (LLMs) necessitates the development of robust detection systems. However, progress

local-aiarxiv-cs-cl
27 Aug 2026
Tutorials

Long-Term Behavioral Evaluation for Trusted Collaborator Selection via Bidirectional Mamba

DGX agent

arXiv:2608.25232v1 Announce Type: new Abstract: Effective selection of trustworthy collaborators is crucial to ensuring the successful completion of collaborative tasks, which requires accurate assess

tutorialsarxiv-cs-lg
27 Aug 2026
Safety

Lower-Resource, Higher Scores: Language Bias in LLM Evaluators

DGX agent

arXiv:2607.14480v3 Announce Type: replace Abstract: LLM evaluators (trained reward models and prompted LLM-as-a-Judge) are routinely validated via pairwise accuracy. In a multilingual setting, this op

safetyarxiv-cs-cl
27 Aug 2026
Research

Memory-V2V: Memory-Augmented Video-to-Video Diffusion for Consistent Multi-Turn Editing

DGX agent

arXiv:2601.16296v3 Announce Type: replace Abstract: Video-to-video diffusion models achieve impressive single-turn editing performance, but practical editing workflows are inherently iterative. When e

researcharxiv-cs-cv
27 Aug 2026
Model Releases

MMEmb-R1: Reasoning-Enhanced Multimodal Embedding with Pair-Aware Selection and Adaptive Control

DGX agent

arXiv:2604.06156v2 Announce Type: replace-cross Abstract: MLLMs have been successfully applied to multimodal embedding tasks, yet their generative reasoning capabilities remain underutilized. Directly

model-releasesarxiv-cs-cl
27 Aug 2026
Model Releases

Non-Asymptotic Bounds for Closed-Loop Identification of Sub-Exponentially Growing Nonlinear Stochastic Systems

DGX agent

arXiv:2412.04157v2 Announce Type: replace-cross Abstract: We investigate the problem of least squares parameter estimation from single-trajectory data for discrete-time, unstable, closed-loop nonlinea

model-releasesarxiv-cs-lg
27 Aug 2026
Model Releases

ONNX-Net: Towards Universal Representations and Instant Performance Prediction for Neural Architectures

DGX agent

arXiv:2510.04938v2 Announce Type: replace-cross Abstract: Neural architecture search (NAS) automates the design process of high-performing architectures, but remains bottlenecked by expensive performa

model-releasesarxiv-cs-cl
27 Aug 2026
Tutorials

PaSta: Noisy Node Classification with Partial Label Learning

DGX agent

arXiv:2608.25365v1 Announce Type: new Abstract: Noisy node classification problem is a fundamental yet challenging task for real-world graph-related web services, where node labels are often corrupted

tutorialsarxiv-cs-lg
27 Aug 2026
Model Releases

Plans You Can Check: Verifier-Grounded Learning of an Open-Weight Planner for Executable Video-Editing

DGX agent

arXiv:2608.25622v1 Announce Type: cross Abstract: Practical video editing is not only pixel generation: an editor must turn a brief, a clip pool, music metadata, and hard constraints into an executabl

model-releasesarxiv-cs-cl
27 Aug 2026
Model Releases

PlanSightRAG: A Visual-First Multimodal RAG for Automating Question Answering and Compliance Checking for Civil Standard Plans

DGX agent

arXiv:2608.26091v1 Announce Type: cross Abstract: Civil infrastructure compliance checking has long relied on engineers manually reading legacy 2D plans; however, OCR-based automation strips away the

model-releasesarxiv-cs-cl
27 Aug 2026
Safety

PreResQ-R1: Response-Preference Disentangled Ranking-and-Scoring Reinforcement Optimization for Robust Visual Quality Assessment

DGX agent

arXiv:2511.05393v2 Announce Type: replace Abstract: Visual Quality Assessment (QA) seeks to predict human perceptual judgments of visual fidelity. While recent multimodal large language models (MLLMs)

safetyarxiv-cs-cv
27 Aug 2026
Model Releases

PRISM: Projection-Integrated Sampling-Based MPC with Bayesian Cost Tuning for Bimanual Manipulation

DGX agent

arXiv:2608.25666v1 Announce Type: new Abstract: Bimanual manipulation in cluttered, contact-rich environments remains challenging because it requires coordinated motion generation, interaction-aware p

model-releasesarxiv-cs-ro
27 Aug 2026
Model Releases

Qwen3.8-Flash-Next: Time to Update Those Benchmarks

DGX agent

specs hardware: M4 Max 128GB Studio inference engine: oMLX & lllama.cpp insights it still very early, so had to disable oMLX K/V caching, qwen4_exp architectureis not yet supported + the obvious n-gra

model-releasesr-localllama
27 Aug 2026
Tutorials

R^3: Training Robots to Reason in Natural Language via Reinforcement Learning

DGX agent

arXiv:2608.26053v1 Announce Type: cross Abstract: Reasoning in language allows foundation models to spend more test-time compute on hard problems, such as those requiring decomposition, constraint tra

tutorialsarxiv-cs-cl
27 Aug 2026
Applications

Recurrent Neural Networks in Linguistic Theory: Revisiting Pinker and Prince (1988) and the Past Tense Debate

DGX agent

arXiv:1807.04783v3 Announce Type: replace Abstract: Can advances in NLP help advance cognitive modeling? We examine the role of artificial neural networks, the current state of the art in many common

applicationsarxiv-cs-cl
27 Aug 2026
Model Releases

RefLAM: A Reference-Grounded Line Annotation Pipeline for Historical Arabic Manuscripts

DGX agent

arXiv:2608.25140v1 Announce Type: cross Abstract: Existing approaches to building line-level Arabic handwritten-text-recognition (HTR) training data either rely on fully manual annotation, which does

model-releasesarxiv-cs-cl
27 Aug 2026
Safety

Retrieved But Not Reliable: A Survey on Attacks, and Defenses in Retrieval-Augmented Generation

DGX agent

arXiv:2608.24977v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) enhances large language models by grounding outputs in external knowledge, improving factuality and reducing hall

safetyarxiv-cs-cl
27 Aug 2026
Model Releases

Sample Margin-Aware Recalibration of Temperature Scaling

DGX agent

arXiv:2506.23492v2 Announce Type: replace-cross Abstract: Recent advances in deep learning have significantly improved predictive accuracy. However, modern neural networks remain systematically overco

model-releasesarxiv-cs-cv
27 Aug 2026
Agents

Scalable Supervision for Software Agents via Patch Reasoning

DGX agent

arXiv:2510.22775v2 Announce Type: replace Abstract: While language model agents have advanced software engineering, existing test-based supervision is limiting its scalability on real-world issues. Th

agentsarxiv-cs-cl
27 Aug 2026
Model Releases

StablePDENet: Enhancing Neural Operator Stability through Physics-Informed Residual-Sensitivity Regularization

DGX agent

arXiv:2601.06472v2 Announce Type: replace Abstract: Learning solution operators for differential equations with neural networks has shown great potential in scientific computing, but ensuring their st

model-releasesarxiv-cs-lg
27 Aug 2026
Agents

TAU-Agent: An Agentic Retrieval-Augmented Framework for Traffic Anomaly Understanding

DGX agent

arXiv:2608.25935v1 Announce Type: new Abstract: Traffic Anomaly Understanding (TAU) requires models and systems to detect, reason about, and explain anomalous events in transportation videos. To addre

agentsarxiv-cs-cv
27 Aug 2026
Model Releases

The Changing Geometry of Grammar: Dimensionality and Neighborhood Reorganization across Transformer Layers

DGX agent

arXiv:2608.25166v1 Announce Type: new Abstract: Transformer representations describe trajectories through high-dimensional vector spaces, which are shaped dynamically as tokens incorporate relational

model-releasesarxiv-cs-cl
27 Aug 2026
Agents

TraceML: An Empirical Analysis of Human-Agent Planning in Machine Learning Development

DGX agent

arXiv:2608.26086v1 Announce Type: new Abstract: Large language models write correct code for isolated problems but remain far weaker at autonomous machine-learning development, where an agent must rev

agentsarxiv-cs-lg
27 Aug 2026
Model Releases

Trust the Mass: Forced Weights in KV-Cache Eviction

DGX agent

arXiv:2608.25230v1 Announce Type: cross Abstract: Every deployed sparse-attention or KV-cache-eviction rule keeps a subset of the keys, discards the rest, and renormalizes the attention weights over t

model-releasesarxiv-cs-cl
27 Aug 2026
Model Releases

Unfolding Scientific Papers into Multi-Turn Generation Trajectories for Continued Pre-Training

DGX agent

arXiv:2608.25826v1 Announce Type: new Abstract: A recent line of synthetic-data work reconstructs the thinking behind existing text rather than rewriting the text itself, but it operates on short web

model-releasesarxiv-cs-cl
27 Aug 2026
Research

When Personality Meets Quantization: A Layer-wise MBTI Analysis of Quantized LLMs

DGX agent

arXiv:2608.25977v1 Announce Type: new Abstract: Personality is increasingly important in large language models (LLMs), as it shapes users' trust, engagement, and emotional experiences. While the Myers

researcharxiv-cs-cl
27 Aug 2026
Model Releases

When Should a Network Emit Geometry, and When Should It Detect It? Readout, Reconciliation, and Representation in Floorplan Vectorization

DGX agent

arXiv:2608.25608v1 Announce Type: new Abstract: A network trained to recover the walls, openings, and rooms of a rasterized floorplan can produce its output in two ways: by emitting the geometry as an

model-releasesarxiv-cs-cv
27 Aug 2026
Model Releases

40 defensive security-audit skills distilled from Ox Alpha

DGX agent

Recently, I've been building a project where sandboxing and isolation of certain features was required. However, for the life of me, I couldn't ask Fable 5 a simple question without it accusing me of

model-releasesr-ollama
26 Aug 2026
Model Releases

Adaptive Influence Graphs for Failure Attribution in Multi-Agent Systems

DGX agent

arXiv:2608.24361v1 Announce Type: new Abstract: Multi-agent LLM systems are increasingly deployed in real-world applications, where failures can be costly and difficult to localize. Despite growing ef

model-releasesarxiv-cs-ai
26 Aug 2026
Model Releases

Adaptive prediction theory combining offline and online learning

DGX agent

arXiv:2512.00342v2 Announce Type: replace Abstract: Real-world intelligence systems usually operate by combining offline learning and online adaptation with highly correlated and non-stationary system

model-releasesarxiv-cs-lg
26 Aug 2026
Model Releases

AI assistant Instinct is raising a 250M Series B co-led by Index Ventures and Benchmark at a 2.5B valuation, bringing its total raised to $350M (Kate Clark/Wall Street Journal)

DGX agent

Kate Clark / Wall Street Journal: AI assistant Instinct is raising a 250M Series B co-led by Index Ventures and Benchmark at a 2.5B valuation, bringing its total raised to 350M — Buzzy startup Instinc

model-releasestechmeme
26 Aug 2026
Local Ai

ALPHABET: A Laplace-Pole History Aggregator with Banked Exponential Transport

DGX agent

arXiv:2608.24051v1 Announce Type: new Abstract: Can a sequence model remain competitive with only a few thousand parameters and an explicitly auditable prediction interface? We introduce ALPHABET, a c

local-aiarxiv-cs-lg
26 Aug 2026
Model Releases

Anyone else doing eGPUs (OCuLink)?

DGX agent

Upgraded to a 5070 Ti so I could run Qwen 3.8 27B, which works perfectly, but didn't want to let the old 4070 Ti go to waste. The cards would touch if I put them both in the PC and I knew the heat wou

model-releasesr-localllama
26 Aug 2026
Model Releases

Are Android GUI Agents Robust Against Runtime Anomalies? AnTrap: Evaluating Agents in Dynamic Adversarial Environments

DGX agent

arXiv:2608.24099v1 Announce Type: new Abstract: GUI agents often encounter dynamic anomalies when deployed on Android devices, from unexpected pop-ups to action misuse, yet existing benchmarks lack sy

model-releasesarxiv-cs-ai
26 Aug 2026
Model Releases

Compared Qwen 3.8 27B community quants on RTX 6000 vs Claude Opus 4.6

DGX agent

*part 2 of an earlier post: previous quant comparison with voxel island creation this time I rented three rtx pro 6000 96gb, on each one I launched a qwen 3.8 27b quant and gave them 4 identical promp

model-releasesr-localllama
26 Aug 2026
Model Releases

Coronavirus Optimization Algorithm: A Success-History Adaptive Evolutionary Framework with Archive-Assisted Search and Stagnation Recovery for Global Optimization

DGX agent

arXiv:2608.23847v1 Announce Type: cross Abstract: This paper proposes the Coronavirus Optimization Algorithm (COA), a SARS-CoV-2-inspired success-history adaptive evolutionary optimizer for box-constr

model-releasesarxiv-cs-ai
26 Aug 2026
Model Releases

Data Predictability Shapes Weibull Weight-Scale Growth in Transformer Training

DGX agent

arXiv:2608.23573v1 Announce Type: new Abstract: A trained transformer's weight magnitudes can be summarized by a two-parameter Weibull distribution whose shape k approx 1.2 is stable across layers and

model-releasesarxiv-cs-lg
26 Aug 2026
Research

Deep Learning Super Resolution for Satellite Cloud Mask Downscaling

DGX agent

arXiv:2608.24715v1 Announce Type: cross Abstract: A vast amount of optical satellite data is being transmitted to Earth-based servers every day, and more than half of this data is affected by haze or

researcharxiv-cs-ai
26 Aug 2026
Research

DriftAD: Visually-Guided Text Drift for Few-Shot Industrial Anomaly Detection

DGX agent

arXiv:2608.23723v1 Announce Type: new Abstract: Few-shot anomaly detection (FSAD) has recently benefited from vision-language models such as CLIP, which enable anomaly de?tection by aligning visual fe

researcharxiv-cs-cv
26 Aug 2026
← Previous
1…576577578579580…1395
Next →