AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,507 results
Model Releases

Revisiting Predictive Process Monitoring in the Age of Foundation Models: A Comparative Study of Sequence, Tabular, and LLM Approaches

DGX agent

arXiv:2607.27797v1 Announce Type: new Abstract: Predictive process monitoring (PPM) leverages event logs to forecast the future of running process instances, for instance, predicting the next activity

model-releasesarxiv-cs-lg
31 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

RMBench: Memory-Dependent Robotic Manipulation Benchmark with Insights into Policy Design

DGX agent

arXiv:2603.01229v3 Announce Type: replace Abstract: Robotic manipulation policies have made rapid progress in recent years, yet most existing approaches give limited consideration to memory capabiliti

model-releasesarxiv-cs-ro
31 Jul 2026
Model Releases

Safety-Gated Agentic Supervisory Control on a Coupled Distillation Benchmark: Regime Map, Auditable Gate, and Co-Design Findings

DGX agent

arXiv:2607.27849v1 Announce Type: cross Abstract: An open-weight LLM can write composition setpoints every five minutes. What a plant still needs is a hard check: named constraints, logged margins, an

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

Same Facts, Different Diagnosis: Measuring and Mitigating Narrative Anchoring in Clinical Language Models

DGX agent

arXiv:2607.27384v1 Announce Type: new Abstract: Large language models used for clinical diagnostic reasoning are sensitive to sociolinguistic register, not just clinical content. We term this failure

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

SARC-DQ: Runtime Data-Quality Gating for Agentic AI: Silent Evidence Defects, the Incompetence Shield, and Downstream-Only Remediation

DGX agent

arXiv:2607.26313v1 Announce Type: cross Abstract: Agentic systems act, so a defect in the evidence they retrieve becomes a wrong action with a currency cost. The most dangerous enterprise defects are

model-releasesarxiv-cs-ai
31 Jul 2026
Model Releases

Scaling medical imaging report generation with multimodal reinforcement learning

DGX agent

arXiv:2601.17151v2 Announce Type: replace-cross Abstract: Frontier models have demonstrated remarkable capabilities in understanding and reasoning with natural-language text, but they still exhibit ma

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

Scaling Vision-Language Models Is Not Enough to Mitigate Bias

DGX agent

arXiv:2607.28211v1 Announce Type: new Abstract: Vision-Language Models (VLMs) such as CLIP are now foundational to multimodal systems, yet their robustness to spurious correlations remains poorly unde

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

Select or Project? Evaluating Lower-dimensional Vectors for LLM Training Data Explanations

DGX agent

arXiv:2601.16651v3 Announce Type: replace Abstract: Gradient-based methods for instance-based explanation for large language models (LLMs) are hindered by the immense dimensionality of model gradients

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

Selecting Open-Weight Language Models for Zero-Shot Intent Classification: A Systematic Evaluation of 41 Models

DGX agent

arXiv:2607.27421v1 Announce Type: new Abstract: Intent classification is a core component of task-oriented dialogue systems, yet practitioners have limited systematic guidance for selecting deployable

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

SemAnCorr: Semantic Anchored Correspondence for Zero-Shot Manipulation Skill Transfer

DGX agent

arXiv:2607.28382v1 Announce Type: new Abstract: Transferring manipulation skills across object instances that share functionality but differ in geometry remains a fundamental challenge in robot learni

model-releasesarxiv-cs-ro
31 Jul 2026
Model Releases

SenseNova U1.5 Lite preview just dropped

DGX agent

SenseNova released U1.5-Lite-Preview Benchmarks: Qwen-Image-Bench from 47.14 to 55.20. ImgEdit-Bench from 3.90 to 4.37. GEdit-Bench-en from 7.47 to 8.17. Key updates: 4K native generation with better

model-releasesr-localllama
31 Jul 2026
Model Releases

Shared Symbolic Backbones for Physically Consistent Multi-Output Symbolic Regression

DGX agent

arXiv:2607.26528v1 Announce Type: cross Abstract: Symbolic regression provides analytical expressions, but it is usually applied one output at a time. This is limiting in process systems, where state

model-releasesarxiv-cs-ai
31 Jul 2026
Model Releases

Sign Language Question Answering: A New Task, Benchmark, and Baseline for Sign Language Understanding

DGX agent

arXiv:2607.27826v1 Announce Type: cross Abstract: Recent advances in sign language (SL) understanding (SLU) have led to remarkable progress in tasks such as continuous SL recognition and SL translatio

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

Simplifying Neural Networks During Training

DGX agent

arXiv:2607.27854v1 Announce Type: cross Abstract: Understanding and exploiting the training dynamics of overparameterized deep neural networks remains a central challenge in modern machine learning. R

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

smevals - a small eval suite for evaluating models, prompts, and harnesses

DGX agent

smevals - a small eval suite for evaluating models, prompts, and harnesses I've been working with Jesse Vincent's Prime Radiant applied AI research lab building out this evals framework to help answer

model-releasessimon-willison
31 Jul 2026
Model Releases

Some deepseek-v4-flash 20260731 opinion review

DGX agent

First of all, I want to apologize if it's off-topic or in the wrong format. Having tried Deepseek Flash with reasoning high on a conceptually difficult task, involving Machine Learning classifiers and

model-releasesr-localllama
31 Jul 2026
Model Releases

Sources: OpenAI demoed a new 'Astra' AI model family to US policymakers and regulators this week, touting its improved abilities to complete long-running tasks (The Information)

DGX agent

The Information: Sources: OpenAI demoed a new “Astra” AI model family to US policymakers and regulators this week, touting its improved abilities to complete long-running tasks — OpenAI is preparing t

model-releasestechmeme
31 Jul 2026
Model Releases

Space2Ground 2.0: A Multi-Source Dataset and Framework for Agricultural Monitoring through Fusion of Street-Level and Satellite Imagery

DGX agent

arXiv:2607.28247v1 Announce Type: new Abstract: Accurate and scalable parcel-level agricultural monitoring remains challenging because satellite Earth Observation alone provides only an overhead persp

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

Stateless MCP has recaptured my interest (and inspired mcp-explorer and datasette-mcp)

DGX agent

Tuesday was Stateless MCP day - the rollout of MCP 2.0, or the 2026-07-28 Model Context Protocol specification to use the more formal but less memorable name. This is the most significant change to th

model-releasessimon-willison
31 Jul 2026
Model Releases

StealthBench: Measuring Operational Stealth in Autonomous Offensive-Security Agents

DGX agent

arXiv:2607.26314v1 Announce Type: cross Abstract: Stealth, the discipline of achieving an objective without revealing your presence, capabilities, or collected intelligence, is what separates sophisti

model-releasesarxiv-cs-ai
31 Jul 2026
Model Releases

STEREODISCO: Discovering Stereotypicality in LLMs

DGX agent

arXiv:2607.27824v1 Announce Type: cross Abstract: LLMs encode, convey, and perpetuate stereotypes. Prior computational research focuses on a small set of semantic axes investigated in social psycholog

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

Subtract or Replay? Exact Deletion from Language-Model Memory

DGX agent

arXiv:2607.27539v1 Announce Type: cross Abstract: Exact deletion from persistent language-model memory depends on how that memory represents a record. Addressable influence can be removed by algebraic

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

Sympathetic Framing: Evaluating AI Alignment across Sociodemographic Groups

DGX agent

arXiv:2607.27232v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly shaping how we consume information and form our worldview. This raises concerns beyond bias in AI: do LLMs

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

TEA-AgriVLN: Traversability Estimation Alarm for Agricultural Vision-and-Language Navigation

DGX agent

arXiv:2607.28474v1 Announce Type: new Abstract: Vision-and-Language Navigation in Continuous Environments (VLN-CE) requires an agent to follow a natural language instruction, predicting a sequence of

model-releasesarxiv-cs-ro
31 Jul 2026
Model Releases

The Convergence Behavior of Adam under Heavy-Tailed Noise

DGX agent

arXiv:2607.27383v1 Announce Type: new Abstract: We establish the first convergence guarantees for the plain vector-form Adam optimizer under heavy-tailed stochastic noise. While several Adam variants

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

The new DeepSeek v4 Flash is now available in Hermes Agent through Nous Portal and OpenRouter

DGX agent

The new DeepSeek v4 Flash is now available in Hermes Agent through Nous Portal and OpenRouter 🚀 DeepSeek-V4-Flash Official API is now LIVE in public beta! 🔷 We’ve massively upgraded its Agent capabili

model-releasesnous-research--x
31 Jul 2026
Model Releases

The price wars that I have predicted going back to August 2023 are now in full swing. All this was inevitable – as long as everyone builds m…

DGX agent

The price wars that I have predicted going back to August 2023 are now in full swing. All this was inevitable – as long as everyone builds more or less the same kind of AI, there is no moat, and profi

model-releasesgary-marcus--x
31 Jul 2026
Model Releases

They knew the charges were bullshit from the start. They knew they arrested an innocent man just to intimidate others and appease the fake n…

DGX agent

They knew the charges were bullshit from the start. They knew they arrested an innocent man just to intimidate others and appease the fake narrative of an incompetent fucking toddler. This is the real

model-releasesanthropic--x
31 Jul 2026
Model Releases

THGFM: Dual-Branch Temporal Heterogeneous Graph Fusion Model

DGX agent

arXiv:2607.27303v1 Announce Type: cross Abstract: Temporal heterogeneous graphs offer a natural abstraction for dynamic relational systems in which diverse node and relation types co-exist and evolve

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

Think with Extra-Image: A Farmland Segmentation Agent Driven by Spatio-Temporal Information Gain

DGX agent

arXiv:2607.28186v1 Announce Type: new Abstract: Existing farmland remote sensing image (FRSI) segmentation follows a 'Think with Intra-Image' paradigm, assuming that the current image contains suffici

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

Thinking Once Is Enough: Intermediate-Layer Evidence Routing for High-Resolution VQA

DGX agent

arXiv:2607.27830v1 Announce Type: new Abstract: High-resolution visual question answering (HR-VQA) is often treated as a problem of insufficient evidence acquisition, where failing multimodal large la

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

This is both a real incident (in that the AI really did get unauthorized access to real systems) and also something it was (sort of) prompte…

DGX agent

This is both a real incident (in that the AI really did get unauthorized access to real systems) and also something it was (sort of) prompted to do. In a review of our cybersecurity evaluations, we fo

model-releasesethan-mollick--x
31 Jul 2026
Model Releases

This will happen frequently as AI becomes smarter and more agentic

DGX agent

This will happen frequently as AI becomes smarter and more agentic In a review of our cybersecurity evaluations, we found three incidents in which a Claude model reached the internet from within or wh

model-releaseselon-musk--x
31 Jul 2026
Model Releases

Tight Sample Complexity for Low-Rank Adaptation: Matching Bounds and Rank Selection

DGX agent

arXiv:2607.27680v1 Announce Type: cross Abstract: Low-Rank Adaptation (LoRA) has become the standard mechanism for fine-tuning large pretrained models, yet its statistical properties remain only parti

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

Towards Generalized Synapse Detection Across Invertebrate Species

DGX agent

arXiv:2509.17041v2 Announce Type: replace Abstract: Behavioural differences across organisms, whether healthy or pathological, are closely tied to the structure of their neural circuits. Yet, the fine

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

Towards Stability of Parameter-Free Optimization

DGX agent

arXiv:2405.04376v4 Announce Type: replace Abstract: Hyperparameter tuning, particularly the selection of an appropriate learning rate in adaptive gradient training methods, remains a challenge. To add

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

Towards Unified Multimodal Misinformation Detection in Social Media: A Benchmark Dataset and Baseline

DGX agent

arXiv:2509.25991v3 Announce Type: replace-cross Abstract: Detecting deceptive multimodal content on social media has become an increasingly important problem. Two major types of deception dominate: hu

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

TraceCoder: Explainable and Auditable Code Generation with Position-Key Snippet Versioning

DGX agent

arXiv:2607.26307v1 Announce Type: new Abstract: Contemporary LLM-based coding agents produce code as black-box outputs: the rationale behind each line is hidden, the evolution of the code through benc

model-releasesarxiv-cs-ai
31 Jul 2026
Model Releases

TriShield: Zero-Utility-Loss Defense Against Privacy Backdoors in Federated Language Model Fine-Tuning via Orthogonal Gradient Projection and Optimizer State Entanglement

DGX agent

arXiv:2607.27940v1 Announce Type: cross Abstract: Federated fine-tuning of large language models (LLMs) enables collaborative training without exposing raw data. However, a recent attack, NeuroImprint

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

Try Grok 4.5 http://X.ai/cli

DGX agent

Try Grok 4.5 http://X.ai/cli BREAKING: Grok 4.5 outperforms GPT-5.6 Terra across almost all shared benchmarks on AskClash, leading in ACB, GPQA, SWE-P, and Atlas while also achieving a higher overall

model-releaseselon-musk--x
31 Jul 2026
Model Releases

Tycho: Active Abstraction with Programmatic World Models for ARC-AGI-3

DGX agent

arXiv:2607.28287v1 Announce Type: cross Abstract: ARC-AGI-3 turns abstraction into an interactive problem of skill acquisition. A player must infer an unfamiliar game's rules, hidden state, and goal w

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

Uncensored Multi-Model Releases, Jamba2-Mini, Qwen3.5-9B-Nikusui-v1 with MTPs and Qwen3.5-27B-Nikusui-v1 with MTPs, Available in Safetensors and GGUF Formats!

DGX agent

First we have Jamba2-Mini Ultra Uncensored Heretic, it's a model which has never been uncensored before, it's a hybrid Mamba model with 52B parameters. Here is the model links: Safetensors: https://hu

model-releasesr-ollama
31 Jul 2026
Model Releases

Uncensored Multi-Model Releases, LongCat-Flash-Lite with MTPs, Jamba2-Mini, Qwen3.5-9B-Nikusui-v1 with MTPs and Qwen3.5-27B-Nikusui-v1 with MTPs, Available in Safetensors and GGUF Formats!

DGX agent

Been working hard for the past month to bring to the community some interesting curios, so for starters we have LongCat-Flash-Lite Uncensored Heretic with MTPs which has never before been uncensored,

model-releasesr-localllama
31 Jul 2026
Model Releases

UrbanDS: A Graph-Guided LLM Multi-Agent System for Data-Intensive Urban Tasks

DGX agent

arXiv:2607.26724v1 Announce Type: new Abstract: Large language model (LLM) agents have been widely applied in automating data science tasks. However, existing methods typically rely on a limited set o

model-releasesarxiv-cs-ai
31 Jul 2026
Model Releases

Using an AMD V620 workstation card for ComfyUI - success

DGX agent

A few weeks ago I posted about if it was worth using a V620 for Comfyui, and was told it likely wouldn't work, at least in Windows 11. And if it did, it would be far too slow and unusable. I decided t

model-releasesr-stablediffusion
31 Jul 2026
Model Releases

Using Large Language Models for Idea Generation in Innovation

DGX agent

arXiv:2607.27553v1 Announce Type: cross Abstract: This research evaluates the efficacy of large language models (LLMs) in generating new product ideas. To do so, we compare three pools of ideas for ne

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

Very interesting paper on recursive self-improvement. The whole stack is released. Machine learning engineering gives recursive self-improve…

DGX agent

Very interesting paper on recursive self-improvement. The whole stack is released. Machine learning engineering gives recursive self-improvement a concrete, executable testbed. OpenMLE is an open full

model-releasesdair-ai--x
31 Jul 2026
Model Releases

VESTIGE: A Knowledge-Guided Masking Strategy for Corruption-Aware Fine-Tuning of Genomic Transformers, Validated on Ancient DNA Reconstruction

DGX agent

arXiv:2607.27712v1 Announce Type: new Abstract: Standard masked-language-model fine-tuning applies a uniform masking probability across every token position, assuming reconstruction difficulty is posi

model-releasesarxiv-cs-lg
31 Jul 2026
← Previous
1…6465666768…469
Next →