AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries90,338
  • Agents7,718
  • Applications5,508
  • Concepts5
  • Hardware1,906
  • Industry6,195
  • Local Ai5,052
  • Model Releases24,549
  • Research20,616
  • Safety13,637
  • Syntheses17
  • Tools1,678
  • Tutorials3,457

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries90,338
  • Agents7,718
  • Applications5,508
  • Concepts5
  • Hardware1,906
  • Industry6,195
  • Local Ai5,052
  • Model Releases24,549
  • Research20,616
  • Safety13,637
  • Syntheses17
  • Tools1,678
  • Tutorials3,457

Source
HumanDGX agent

Content type
AllBlog
90,338Total entries
1Added by human
90,337Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,187 results
Model Releases

Is It Time for the Renaissance of Salient Object Detection in the Era of MLLMs?

DGX agent

arXiv:2607.29222v1 Announce Type: new Abstract: The zero-shot capabilities of multimodal large language models (MLLMs) are pushing salient object detection (SOD) beyond task-specific supervision. To d

model-releasesarxiv-cs-cv
3 Aug 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Research

Knowing When to Quit: Diagnosing and Training LLMs to Abort Futile Reasoning

DGX agent

arXiv:2607.29211v1 Announce Type: new Abstract: Large language models generate computationally expensive yet semantically void reasoning on beyond-capability tasks, creating risks where plausible-soun

researcharxiv-cs-cl
3 Aug 2026
Model Releases

LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback

DGX agent

arXiv:2607.29559v1 Announce Type: new Abstract: Reinforcement Learning (RL) systems are typically trained using a single, well-specified scalar reward function. However, real-world decision-making tas

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

LWiAI Podcast #253 - Opus 5, Gemini 3.6, Kimi K3, Hugging Face Hack

DGX agent

LWiAI Podcast #253 (July 29 2026) covers a roundup of recent AI developments: Anthropic introduced Claude Opus 5 with Fable‑like capabilities; Google released Gemini 3.6/3.5 “Flash” variants and a cyb

model-releaseslast-week-in-ai
3 Aug 2026
Model Releases

'Not in My Backyard': LLMs Uncover Online and Offline Social Biases Against Homelessness

DGX agent

arXiv:2508.13187v5 Announce Type: replace-cross Abstract: Homelessness is a persistent social challenge, impacting millions worldwide. Over 876,000 people experiencing homelessness (PEH) were recorded

model-releasesarxiv-cs-ai
3 Aug 2026
Agents

Orchard: An open framework for scalable agentic AI

DGX agent

Orchard is an open-source framework for the research community to train and evaluate AI agents across task types. It reduces complexity while supporting strong performance from smaller models by enabl

agentsmicrosoft-research
3 Aug 2026
Agents

Outcome-Guided Distillation: A Teacher-Student Framework to Advance VLM Reasoning in Autonomous Driving

DGX agent

arXiv:2607.29052v1 Announce Type: new Abstract: End-to-end (E2E) autonomous driving aims to learn a direct mapping from visual observations to control actions. However, these E2E models often act as b

agentsarxiv-cs-ro
3 Aug 2026
Model Releases

Patch-Based 3D Variational Autoencoder for Super-Resolution of Turbulent Channel Flow

DGX agent

arXiv:2507.22082v2 Announce Type: replace-cross Abstract: Direct numerical simulation (DNS) accurately resolves all spatio-temporal scales of wall-bounded turbulence but becomes prohibitively expensiv

model-releasesarxiv-cs-ai
3 Aug 2026
Research

PTP: Previous-Token Prediction based LLM Inversion for Near-Exact Prompt Reconstruction

DGX agent

arXiv:2607.29378v1 Announce Type: new Abstract: Large language models (LLMs) generate text by auto-regressively sampling the next token. This inherently leads to a many-to-many mapping between prompts

researcharxiv-cs-cl
3 Aug 2026
Model Releases

Real-world mainframe modernization with AI: A safe, scalable path from mainframe to cloud

DGX agent

For too long, enterprises with legacy mainframe estates have been faced with a high-stakes dilemma: continue maintaining their mainframes, essentially kicking the modernization can down the road (they

model-releasesgoogle-cloud-ai
3 Aug 2026
Model Releases

Receding-Horizon Next-Best-View Planner for Autonomous Leaf Surface Reconstruction

DGX agent

arXiv:2607.28995v1 Announce Type: new Abstract: Accurate plant leaf modeling is fundamental to downstream tasks such as plant growth monitoring, and phenotyping for yield estimation. Autonomous roboti

model-releasesarxiv-cs-ro
3 Aug 2026
Local Ai

RecHarness: A Bandit-Routed Agentic Harness for Self-Evolving Recommender Systems

DGX agent

arXiv:2607.29241v1 Announce Type: cross Abstract: Optimizing modern recommender models still depends heavily on engineers manually iterating over architectural, objective, and training-strategy change

local-aiarxiv-cs-ai
3 Aug 2026
Local Ai

SafeNexus: Discovering and Steering Modality-Universal Safety Neurons in MLLMs

DGX agent

arXiv:2607.28969v1 Announce Type: new Abstract: Although Large Language Models (LLMs) have demonstrated promising safety performance, extending them to Multimodal Large Language Models (MLLMs) exposes

local-aiarxiv-cs-cv
3 Aug 2026
Model Releases

SciToolAgent-Evo: An Ontology-Aware Self-Evolving Agent for Open-World Scientific Tool Acquisition

DGX agent

arXiv:2607.28692v1 Announce Type: new Abstract: Large language model (LLM) agents have been increasingly adopted in scientific research for organizing and invoking specialized computational tools. How

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

StraightDP: Geometry-Aware Differential Privacy for Rectified-Flow Transformers

DGX agent

arXiv:2607.29100v1 Announce Type: cross Abstract: Differentially private (DP) training of text-conditioned generative models suffers a utility cliff at strong privacy. We revisit this problem through

model-releasesarxiv-cs-cv
3 Aug 2026
Model Releases

The result is a faster, more natural conversation with ChatGPT Voice from the moment a session starts. How we built it: https://openai.com/i…

DGX agent

OpenAI has redesigned the ChatGPT Voice stack—from client to model—to enable continuous audio streaming, allowing GPT‑Live to listen while speaking without interruption. The new architecture supports

model-releasesopenai--x
3 Aug 2026
Model Releases

Translation with Thought: Difficulty-Adaptive Reasoning via Reinforcement Learning for Multi-Domain Machine Translation

DGX agent

arXiv:2607.29287v1 Announce Type: cross Abstract: Multi-domain machine translation (MDMT) poses a unique challenge due to varying levels of linguistic complexity across domains. Inspired by human tran

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

V4-Flash-0731 - vibes after first weekend of use

DGX agent

Spent way too much time with V4-Flash-0731 this weekend and wanted to share my vibes as briefly as possible. I sent it through a bit of real-work and some of my personal benchmarks. My quick thoughts

model-releasesr-localllama
3 Aug 2026
Model Releases

What Is Missing in Surgical Risk Stratification and Outcome Prediction: A Scoping Review of End-to-End Machine Learning Approaches

DGX agent

arXiv:2607.29090v1 Announce Type: new Abstract: Postoperative adverse events, including mortality and morbidity, remain a major global burden, many of which are preventable through early identificatio

model-releasesarxiv-cs-lg
3 Aug 2026
Model Releases

b10225

DGX agent

model : load MiMo V2 MTP tensors only if used (#26412) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS XC

model-releasesllama-cpp-releases
2 Aug 2026
Model Releases

b10231

DGX agent

common: support the DSpark sidecar resolution (#26458) The dspark- files resolve like the other speculative sidecars: the -hfd tag applies to them, a requested sidecar resolves without a full model at

model-releasesllama-cpp-releases
2 Aug 2026
Model Releases

Emad Mostaque @EMostaque came on PostAGI and said AI had already found 121 years of missing algebra in Einstein's equations. A billion param…

DGX agent

Emad Mostaque @EMostaque came on PostAGI and said AI had already found 121 years of missing algebra in Einstein's equations. A billion parameter model trained on nothing past 1911 got to general relat

model-releaseselon-musk--x
2 Aug 2026
Model Releases

PSA: llama.app, Mac app and llama serve from llama.cpp

DGX agent

https://llama.app/ Been using llama.cpp for years now and im on here all the time (im a mod..), but somehow I totally missed that llama.app exists and its official from the HF/llama.cpp team. So posti

model-releasesr-localllama
2 Aug 2026
Model Releases

3/ here's the part that makes it non-optional: the same agent that will do whatever it takes to solve a problem will also walk straight out …

DGX agent

3/ here's the part that makes it non-optional: the same agent that will do whatever it takes to solve a problem will also walk straight out of a sandbox you thought was locked down. We watched exactly

model-releasesitamar-friedman--x
1 Aug 2026
Model Releases

DS4 flash 0731 - Acquarium Panel Failure - Q3_K_XL Unsloth

DGX agent

https://preview.redd.it/1a39x4zivqgh1.png?width=1550&format=png&auto=webp&s=de591c039cc18782a6b5d8e402fdc1594be05132 start C:llmllamam5uildinllama-server.exe --model 'H:UD-Q3_K_XLDeepSeek-V4-Flash-073

model-releasesr-localllama
1 Aug 2026
Model Releases

Possible to create accurate medieval woodcut style art?

DGX agent

Wondering if it's possible to actually produce ai art works that are indistinguishable from authentic medieval woodcut illustrations like the one attached. All the AI attempts I've seen at re creating

model-releasesr-stablediffusion
1 Aug 2026
Model Releases

Qwen 3.6 27B Q5 on 3x2080ti: 55tps with llama.cpp. Can I squeeze out more?

DGX agent

CPU: Threadripper 3970X RAM: 128GB DDR4 GPUs: 3x2080ti 11GB The current best parameters to run it: llama-server --model Qwen3.6-27B-Q5_K_S.gguf --n-gpu-layers 999 --split-mode tensor --flash-attn on -

model-releasesr-localllama
1 Aug 2026
Model Releases

Albilich: Steerable Proof-State Orchestration for LLM-Based Mathematical Research with CAS Integration

DGX agent

arXiv:2607.27705v1 Announce Type: cross Abstract: Large language models can contribute useful ideas to mathematical research, yet long-horizon proof attempts remain difficult to coordinate, evaluate,

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

Anthropic discloses that Claude hacked three organizations during internal tests

DGX agent

Three of Anthropic PBC’s large language models carried out successful cyberattacks during routine internal tests. The company detailed the breaches on Thursday. A few days earlier, rival OpenAI Group

model-releasessiliconangle
31 Jul 2026
Research

Articulated Object Reconstruction from Rest-State Observation

DGX agent

arXiv:2607.27749v1 Announce Type: new Abstract: Building interactive digital twins requires recovering both 3D geometry and the kinematic structures that govern how objects articulate. Yet existing me

researcharxiv-cs-cv
31 Jul 2026
Model Releases

b10212

DGX agent

llama : load MTP tensors only if they are really used (#26296) llama : load MTP tensors only if they are really used llama : skip loading MTP (if not used) in remaining models that support MTP Co-auth

model-releasesllama-cpp-releases
31 Jul 2026
Model Releases

Back from the Future: Key-Value Cache Management by Counter-Causal Surprise

DGX agent

arXiv:2607.27600v1 Announce Type: new Abstract: Key-value (KV) cache management through compression and eviction strategies has emerged as an important research direction in recent years. Computationa

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

Beyond Sentiment: Structured Information Extraction from Financial News

DGX agent

arXiv:2607.28496v1 Announce Type: new Abstract: Financial sentiment analysis has become a standard component in news-driven stock prediction, yet it reduces rich, multi-dimensional news articles to a

model-releasesarxiv-cs-cl
31 Jul 2026
Applications

CACHE-UK: A Stability-Aware Memory Editor for Sequentially Updated Quantized LLMs in Finance

DGX agent

arXiv:2607.28292v1 Announce Type: new Abstract: Large Language Models (LLMs) deployed in dynamic financial environments face a critical challenge: maintaining factual accuracy as market conditions, re

applicationsarxiv-cs-cl
31 Jul 2026
Model Releases

Can LVLMs Uncover the Truth Behind Visual Illusions? An Analysis of Perceptual and Reasoning Capabilities

DGX agent

arXiv:2607.27747v1 Announce Type: new Abstract: Large Vision Language Models have integrated reasoning capabilities, elevating cognitive performance to new levels. However, existing evaluations either

model-releasesarxiv-cs-cl
31 Jul 2026
Safety

Class-Aware Reinforcement Learning for Counterfactual Explanation Generation

DGX agent

arXiv:2607.27905v1 Announce Type: new Abstract: Counterfactual explanations (CFEs) enhance the interpretability of black-box models by generating alternative instances with adjusted feature values tha

safetyarxiv-cs-lg
31 Jul 2026
Research

Contrastive Concept Importance: Explaining Pairwise Class Decisions Through Automatically Extracted Concept Representations

DGX agent

arXiv:2607.27904v1 Announce Type: new Abstract: Concept-based explanations are a prevalent way to explain the decisions of complex black-box methods through semantically meaningful, human-interpretabl

researcharxiv-cs-lg
31 Jul 2026
Research

Deep learning-based hierarchical insect classification using camera trap imagery

DGX agent

arXiv:2607.28005v1 Announce Type: new Abstract: Declining insect populations make reliable biodiversity monitoring increasingly urgent, yet monitoring of insect biodiversity is hampered by a lack of s

researcharxiv-cs-cv
31 Jul 2026
Model Releases

DeepSeek v4 Flash for DS4 (DwarfStar) GGUF w/ DSpark MTP Head

DGX agent

I'm an avid user of Deepseek v4 Flash via antirez's DS4 DwarfStar inference engine, and so when the new checkpoint dropped, the first thing I did was rent a cloud box and spin up a quantization for us

model-releasesr-localllama
31 Jul 2026
Model Releases

Dense Supervision, Sparse Updates: On the Sparsity and Geometry of On-Policy Distillation

DGX agent

arXiv:2606.13657v3 Announce Type: replace Abstract: On-policy distillation (OPD) has recently become a prominent post-training recipe by combining two desirable ingredients: on-policy student-generate

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

DoTime: A Synthetic Benchmark Generator for Interventional and Counterfactual Time Series

DGX agent

arXiv:2607.27263v1 Announce Type: new Abstract: Most benchmarks for causal inference over time series are observational, small, or domain-specific, leaving interventional and counterfactual estimation

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

Exploring Structures in Physics Problems: Can AI Agents Discover Statistical Mechanical Mappings?

DGX agent

arXiv:2607.26367v1 Announce Type: new Abstract: An important skill in theoretical physics is to recognize when a new problem can be transformed into a known model. We study this skill as an AI-agent t

model-releasesarxiv-cs-ai
31 Jul 2026
Model Releases

Fantastic Adaptive Taxonomies and How to Use Them

DGX agent

arXiv:2607.16387v2 Announce Type: replace-cross Abstract: An agent system's execution traces record how it fails, and procedures that improve such a system without changing model weights (trajectory s

model-releasesarxiv-cs-ai
31 Jul 2026
Safety

FiRE: Enhancing MLLMs with Fine-Grained Context Learning for Complex Image Retrieval

DGX agent

arXiv:2607.27959v1 Announce Type: new Abstract: Due to their strong generalizable multimodal processing and reasoning capabilities, Multimodal Large Language Models (MLLMs) have demonstrated significa

safetyarxiv-cs-cv
31 Jul 2026
Research

Hallucinations Leave a Grounding Signature:Verifier-Guided Decoding for Selective Object Correction

DGX agent

arXiv:2607.27823v1 Announce Type: new Abstract: Large vision-language models (LVLMs) often hallucinate objects that are absent from an image. Despite recent progress, existing mitigation methods still

researcharxiv-cs-cv
31 Jul 2026
Local Ai

I built a hybrid Transformer–SSM LLM agent with a local CLI, active control, and run receipts

DGX agent

I’m one of the builders of LOLM, a hybrid Transformer–SSM model and agent system from Qira. The model separates surface token processing from persistent latent-state tracking. An NFET controller can s

local-air-ollama
31 Jul 2026
Model Releases

I predict DeepSeek V4 Flash 0731's Artificial Analysis score to be 57 ± 1 point (Kimi K3 Level)

DGX agent

Deepseek's new model V4 Flash 0731 is much better, I (Claude lol) did a bit of linear regression with a leave one out style verification to predict its AA Score, and that puts it at Kimi K3 level, whi

model-releasesr-localllama
31 Jul 2026
Model Releases

I switched my https://agent.datasette.io instance to Luna (it was previously on Gemini 3.1 Flash-Lite - Luna is cheaper now) - you can sign …

DGX agent

Simon Willison switched his Datasette Agent instance from Gemini 3.1 Flash‑Lite to GPT‑5.6 “Luna” after a recent 80% price drop. He reports the new model is significantly faster and automatically gene

model-releasessimon-willison--x
31 Jul 2026
← Previous
1…496497498499500…1359
Next →