AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,814
  • Agents7,519
  • Applications5,378
  • Concepts5
  • Hardware1,822
  • Industry6,162
  • Local Ai4,908
  • Model Releases23,658
  • Research20,008
  • Safety13,291
  • Syntheses17
  • Tools1,674
  • Tutorials3,372

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,814
  • Agents7,519
  • Applications5,378
  • Concepts5
  • Hardware1,822
  • Industry6,162
  • Local Ai4,908
  • Model Releases23,658
  • Research20,008
  • Safety13,291
  • Syntheses17
  • Tools1,674
  • Tutorials3,372

Source
HumanDGX agent

Content type
AllBlog
87,814Total entries
1Added by human
87,813Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,138 results
Safety

Proteus: A Truncation-Robust Entropy Model for Progressive LiDAR Compression

DGX agent

arXiv:2608.00687v1 Announce Type: new Abstract: LiDAR point clouds provide explicit, deterministic physical boundaries critical for collaborative safety-critical perception. However, wireless channels

safetyarxiv-cs-cv
4 Aug 2026
Model Releases
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Qwen3.8-Max is available in Hermes Agent now! Let's build! 🚀🚀

DGX agent

Qwen 3.8‑Max, Alibaba.Qwen’s latest large‑language model, has been added to Hermes Agent and can currently be accessed at a 20 % discount. The update aims to streamline integration of the model for de

model-releasesqwen--x
4 Aug 2026
Model Releases

Refine Drugs, Don't Complete Them: Uniform-Source Discrete Flows for Fragment-Based Drug Discovery

DGX agent

arXiv:2509.26405v2 Announce Type: replace Abstract: We introduce InVirtuoGen, a discrete flow generative model for fragmented SMILES for de novo and fragment-constrained generation, and target-propert

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

Shieldstral is available under Apache 2.0. Try it: https://huggingface.co/mistralai/Shieldstral-1.0

DGX agent

**Shieldstral** is a 3‑billion‑parameter open‑weights model developed by Mistral AI for content safety. It can be deployed on-device and is distributed under the Apache 2.0 license. The model is avail

model-releasesmistral-ai--x
4 Aug 2026
Model Releases

Similarity-Aware Machine Unlearning

DGX agent

arXiv:2608.00246v1 Announce Type: new Abstract: Machine unlearning removes the influence of user-specified training examples from a trained model, avoiding the need to retrain it from scratch. Localiz

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

Toward Robust LLM-Based Judges: Taxonomic Bias Evaluation and Debiasing Optimization

DGX agent

arXiv:2603.08091v2 Announce Type: replace Abstract: Large language model (LLM)-based judges are widely adopted for automated evaluation and reward modeling, yet their judgments are often affected by j

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

TreeProbe : A Tibetan Medicine Benchmark for Cultural Bias in LLMs

DGX agent

arXiv:2608.00640v1 Announce Type: new Abstract: Large language models are increasingly viewed as a potential means of mitigating global health inequities, yet their outputs often reflect dominant high

model-releasesarxiv-cs-cl
4 Aug 2026
Safety

Who Should Be Generated? Justifying Demographic Targets in Open-Ended Generation

DGX agent

arXiv:2608.02551v1 Announce Type: cross Abstract: Fairness evaluation concerns not only what a model produces, but also what its outputs ought to be compared against. When a model generates 'a CEO in

safetyarxiv-cs-cl
4 Aug 2026
Research

A Generalized-Bayes Perspective on Counterfactual Explanations: Posterior-Based Decision-Making and Evaluation

DGX agent

arXiv:2607.29077v1 Announce Type: new Abstract: Counterfactual explanations (CEs) enhance the interpretability of machine learning models by identifying the smallest change to an input required to obt

researcharxiv-cs-ai
3 Aug 2026
Model Releases

Agentic Harness for Real-World Compilers

DGX agent

arXiv:2603.20075v2 Announce Type: replace-cross Abstract: Compilers are critical to modern computing, yet fixing compiler bugs is difficult. While recent large language model (LLM) advancements enable

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

b10237

DGX agent

llama : MTP support for DeepSeek V3.2 (#26457) llama : MTP support for DeepSeek V3.2 model : no need to include MTP layers during DeepSeek V3.2 model type discovery Co-authored-by: Stanisław Szymczyk

model-releasesllama-cpp-releases
3 Aug 2026
Model Releases

Benchmarks Are Not Monolithic: Sample-Level Auditing and Orchestration for LLM Evaluation

DGX agent

arXiv:2607.28801v1 Announce Type: cross Abstract: Benchmark datasets are central to evaluating Large Language Models (LLMs), yet they are typically conceived as monolithic tasks, obscuring substantial

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

Benchmarks Are Not Validation: A System-Level View of Financial LLM Applications

DGX agent

arXiv:2607.28840v1 Announce Type: new Abstract: Large language models are increasingly deployed in financial applications that combine retrieval, proprietary data, tool use, orchestration logic, monit

model-releasesarxiv-cs-cl
3 Aug 2026
Research

Can Synthetic Data Overcome the Generalization Limits of AI-Based Flower and Pod Detection Across Cowpea Breeding Genotypes and Environments?

DGX agent

arXiv:2607.28796v1 Announce Type: new Abstract: High-throughput phenotyping requires AI-enabled computer vision models that generalize across genotypes, locations, and growing seasons, yet such models

researcharxiv-cs-cv
3 Aug 2026
Agents

HarnessBank: Semantic Gene-Bank Search with Gated Verification for Agent-Harness Self-Evolution

DGX agent

arXiv:2607.13683v2 Announce Type: replace Abstract: Large Language Models (LLMs) have enabled capable agents across diverse applications. Beyond the foundation model, the performance of an agent is go

agentsarxiv-cs-cl
3 Aug 2026
Model Releases

I gave five different local LLMs a town. They invented Facebook and a duck-based credit bureau. (MIT, self-hosted, you don't play it — you watch it)

DGX agent

Each villager in Pepperton is a different model — a mistral, a qwen3, a qwen2.5, a phi4-mini, a llama3.2 — because model families have genuinely different temperaments, and the friction between them i

model-releasesr-ollama
3 Aug 2026
Model Releases

LARA: Lightweight Adapters in the Residual Stream for Composable Adaptation and Alignment

DGX agent

arXiv:2607.28669v1 Announce Type: new Abstract: We present LARA (Lightweight Additive Residual Adaptation), a method for efficient adaptation that operates in the residual stream of a frozen model rat

model-releasesarxiv-cs-lg
3 Aug 2026
Model Releases

🔥Let's talk about Qwen! #AMA

DGX agent

The post announces an “Ask Me Anything” (AMA) about Qwen, the Alibaba‑developed foundation model. It introduces the Qwen Foundation Model Team and directs readers to their GitHub repository @QwenDevs,

model-releasesqwen--x
3 Aug 2026
Local Ai

Multimodal Reinforcement Learning with Adaptive Verifier for AI Agents

DGX agent

arXiv:2512.03438v3 Announce Type: replace Abstract: Agentic reasoning models trained with multimodal reinforcement learning (MMRL) have become increasingly capable, yet they are almost universally opt

local-aiarxiv-cs-ai
3 Aug 2026
Applications

Pay for The Second-Best Service: A Game-Theoretic Approach Against Dishonest LLM Providers

DGX agent

arXiv:2511.00847v5 Announce Type: replace-cross Abstract: The widespread adoption of Large Language Models (LLMs) through Application Programming Interfaces (APIs) induces a critical vulnerability: th

applicationsarxiv-cs-ai
3 Aug 2026
Local Ai

Small Is Enough: Per-User Style Rewriting of AI-Edited Text via LoRA Adapters

DGX agent

arXiv:2607.29238v1 Announce Type: cross Abstract: InMyStyle is a privacy first, single user system that adapts small language models to rewrite AI-edited text towards an individual user's writing styl

local-aiarxiv-cs-ai
3 Aug 2026
Model Releases

So-Fake: Benchmarking and Explaining Social Media Image Forgery Detection

DGX agent

arXiv:2505.18660v5 Announce Type: replace Abstract: Recent advances in AI-powered generative models have enabled the creation of increasingly realistic synthetic images, posing significant risks to in

model-releasesarxiv-cs-cv
3 Aug 2026
Model Releases

WebCoderBench: Benchmarking Web Application Generation with Comprehensive and Interpretable Evaluation Metrics

DGX agent

arXiv:2601.02430v3 Announce Type: replace-cross Abstract: Web applications (web apps) have become a key arena for large language models (LLMs) to demonstrate their code generation capabilities and com

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

Parlor v2: best-effort fully local GPT-Live clone on an M3 Pro

DGX agent

GPT-Live is so good that I use it almost every day. I've been wanting to replicate it since it was released. My first attempt was to fine-tune Gemma 4 12B to behave like a full-duplex model. Something

model-releasesr-localllama
2 Aug 2026
Model Releases

Ran DS V4-Flash-0731 Locally on 3xMI50 32GB @ ~15 t/s TG

DGX agent

Hey y'all. I'll be concise. TL;DR: DS V4-Flash-0731 @ UD-IQ2_M running fully in VRAM on 3xMI50s (90.9 GB model, 96 GB VRAM). Actual speed on llama-server is: - Text Generation: ~15-16 tokens/second st

model-releasesr-localllama
2 Aug 2026
Local Ai

Try handling complex tasks to your local models with GraphARC, graph engineering yes !

DGX agent

🚀 We just built our first real-time implementation of Graph Engineering, inspired by our experience building graph tooling used by 4,000+ developers. 🔗 Repo: https://github.com/CodeGraphContext/grapha

local-air-localllama
2 Aug 2026
Model Releases

Why are almost all new benchmarks and leaderboards coding focused?

DGX agent

I know in in this community LLM's are generally used for coding but there are other usecases besides coding and those usecases should be tested too. I also know benchmarks can sometimes be benchmaxxed

model-releasesr-localllama
2 Aug 2026
Applications

ARES: Anomaly Recognition Model For Edge Streams

DGX agent

arXiv:2511.22078v2 Announce Type: replace Abstract: Many real-world scenarios involving streaming information can be represented as temporal graphs, where data flows through dynamic changes in edges o

applicationsarxiv-cs-lg
31 Jul 2026
Model Releases

Benchmarking LLM Competence on Logical Inference over Probability Operators

DGX agent

arXiv:2607.27405v1 Announce Type: new Abstract: Both expressions of uncertainty and inferences are ubiquitous in natural language, and valid inferences over natural-language expressions of uncertainty

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

Beyond Borrowed Histories: Person-Aligned User Simulation for Interactive Role-Playing Evaluation

DGX agent

arXiv:2607.27816v1 Announce Type: new Abstract: Role-playing agents (RPAs) have become one of the most important consumer applications of large language models. Users engage in multi-turn conversation

model-releasesarxiv-cs-cl
31 Jul 2026
Research

Calibrate Before Reason: Robust Visual Token Reduction against Semantic Drift in VLMs

DGX agent

arXiv:2607.27700v1 Announce Type: new Abstract: Large Vision-Language Models (VLMs) suffer from prohibitive inference overhead due to long sequences of visual tokens. However, existing visual token re

researcharxiv-cs-cv
31 Jul 2026
Applications

Challenges in annotations by humans and LLMs: A case study of evaluative language

DGX agent

arXiv:2607.28119v1 Announce Type: new Abstract: In this paper, we draw a comparison between linguists in training, a trained linguist, and annotations generated by large language models (LLMs) to find

applicationsarxiv-cs-cl
31 Jul 2026
Model Releases

Critical attention scaling in long-context transformers

DGX agent

arXiv:2510.05554v2 Announce Type: replace Abstract: As large language models scale to longer contexts, attention layers suffer from a fundamental pathology: attention scores collapse toward uniformity

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

CXR-Retrieve: Compositional Text-to-Image Retrieval in Chest Radiography

DGX agent

arXiv:2607.27779v1 Announce Type: new Abstract: Large chest radiography archives are difficult to search because most studies are paired only with free-text reports rather than structured clinical ann

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

EHGCN: Hierarchical Euclidean-Hyperbolic Fusion via Motion-Aware GCN for Hybrid Event Stream Perception

DGX agent

arXiv:2504.16616v4 Announce Type: replace Abstract: Event cameras, characterized by microsecond temporal resolution and very High Dynamic Range (HDR), emit high-speed event streams for perception task

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

eta-OPSD: Deriving with Policy Optimization, Training with Self-Distillation

DGX agent

arXiv:2607.28582v1 Announce Type: new Abstract: On-policy self-distillation (OPSD) is a promising approach to improve reasoning language models, but it remains brittle in practice: making it work reli

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

Exposure is not manifestation: measurement target and output resolution jointly determine which behavioural-faithfulness evaluator wins

DGX agent

arXiv:2607.09306v3 Announce Type: replace Abstract: Behavioural auditing asks whether a language model behaves as it claims, but detection scores are reported without separating two targets: whether a

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

Fairness Pruning: Locating Demographic Bias in GLU-MLP Layers via Differential Activations

DGX agent

arXiv:2607.28319v1 Announce Type: new Abstract: This work presents Fairness Pruning, a lightweight structural intervention method designed for the management and future mitigation of demographic bias

model-releasesarxiv-cs-cl
31 Jul 2026
Safety

Fidelity Is Not Safety: Gently-Compressed LLMs Pass Every Data-Free Quality Guard Yet Invent Procedure Steps in Agentic Execution

DGX agent

arXiv:2607.28196v1 Announce Type: new Abstract: Practitioners accept a compressed language model once it clears a stack of data-cheap quality guards: perplexity within a small factor of the original,

safetyarxiv-cs-cl
31 Jul 2026
Model Releases

Kalman Meets Curriculum: Efficient Dynamic Prompt Selection for Adaptive RL Finetuning

DGX agent

arXiv:2607.27610v1 Announce Type: new Abstract: Reinforcement learning (RL) finetuning significantly enhances the reasoning capabilities of large language models (LLMs), yet its effectiveness critical

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

Language Diversity: Evaluating Language Usage and AI Performance on African Languages in Digital Spaces

DGX agent

arXiv:2512.01557v3 Announce Type: replace Abstract: This study examines the digital representation of African languages and the challenges this presents for current language detection tools. We evalua

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

LLM Self-Correction with DeCRIM: Decompose, Critique, and Refine for Enhanced Following of Instructions with Multiple Constraints

DGX agent

arXiv:2410.06458v2 Announce Type: replace Abstract: Instruction following is a key capability for LLMs. However, recent studies have shown that LLMs often struggle with instructions containing multipl

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

LoMeVQA: A Comprehensive Benchmark for Longitudinal Medical VQA

DGX agent

arXiv:2607.27806v1 Announce Type: new Abstract: In clinical practice, patients often undergo multiple imaging examinations over successive visits, yielding longitudinal data. Modeling such temporal in

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

MPIE-Bench: Benchmarking Anatomically Plausible Multi-Person Interaction Editing

DGX agent

arXiv:2607.27616v1 Announce Type: new Abstract: Text-to-image and personalized editing models now synthesize high-fidelity single-subject images with ease. Yet placing multiple named people into share

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

MSCM-net: A hyperspectral image classiffcation method based on multi-scale convolution and Mamba

DGX agent

arXiv:2607.28277v1 Announce Type: new Abstract: Hyperspectral imaging is widely used in remote sensing and engineering. Therefore, research on its classification methods is crucial. While CNN and Tran

model-releasesarxiv-cs-cv
31 Jul 2026
Tutorials

Noisy Data is Destructive to Reinforcement Learning with Verifiable Rewards

DGX agent

arXiv:2603.16140v2 Announce Type: replace Abstract: Reinforcement learning with verifiable rewards (RLVR) has driven recent capability advances of large language models across various domains. Recent

tutorialsarxiv-cs-lg
31 Jul 2026
Research

S-CEReBrO: Breaking the Memory Barrier in Continuous EEG Monitoring

DGX agent

arXiv:2607.27913v1 Announce Type: new Abstract: Foundation models offer a promising paradigm for Electroencephalography (EEG) analysis, leveraging generalizable representations from vast unlabeled dat

researcharxiv-cs-lg
31 Jul 2026
Model Releases

Sign Language Question Answering: A New Task, Benchmark, and Baseline for Sign Language Understanding

DGX agent

arXiv:2607.27826v1 Announce Type: cross Abstract: Recent advances in sign language (SL) understanding (SLU) have led to remarkable progress in tasks such as continuous SL recognition and SL translatio

model-releasesarxiv-cs-cv
31 Jul 2026
← Previous
1…347348349350351…1316
Next →