AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
All
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,280 results
Model Releases

ChainClaw: A Layered Agent Framework for Reliable On-Chain Execution

DGX agent

arXiv:2608.05790v1 Announce Type: new Abstract: General-purpose large language model agents have achieved strong performance on tool-augmented tasks, yet they rely on assumptions break down in blockch

model-releasesarxiv-cs-ai
7 Aug 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

ChronoVision: Temporal Reasoning via Latent State Reconstruction

DGX agent

arXiv:2608.05631v1 Announce Type: new Abstract: Multimodal large language models excel at passive perception but struggle with complex visual cognitive tasks requiring multi-step temporal reasoning. T

model-releasesarxiv-cs-cv
7 Aug 2026
Model Releases

CLARA: Clarification of Language Ambiguity through Result Analysis for Natural-Language Cancer Genomics Queries

DGX agent

arXiv:2608.05195v1 Announce Type: cross Abstract: A natural language interface can be used to make cancer genomics databases easier to use, but even if a question is perfectly fluent, its scientific m

model-releasesarxiv-cs-lg
7 Aug 2026
Model Releases

Clinician input steers AI toward accurate and harmful recommendations

DGX agent

arXiv:2603.14158v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are entering clinical workflows, yet evaluations rarely assess how clinician reasoning shapes model behavior duri

model-releasesarxiv-cs-lg
7 Aug 2026
Model Releases

CNM-BERT: A Drop-In Structural Embedding for Chinese Characters via Ideographic Description Sequences

DGX agent

arXiv:2608.05167v1 Announce Type: new Abstract: Token-based encoders like BERT treat Chinese characters as atomic identifiers, ignoring their recursive orthographic structure. Consequently, models rel

model-releasesarxiv-cs-cl
7 Aug 2026
Model Releases

CodeGrep: An RL-Trained Retrieval Agent for LLM Coding Agents

DGX agent

arXiv:2608.05886v1 Announce Type: cross Abstract: Modern LLM coding agents such as Claude Code and OpenHands share a common inefficiency: they spend much of their token budget finding the file to patc

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

Codex is overoptimised for large models: it ranks 2nd out of 10 for GLM 5.2 but drops to 9th place for Gemma-4! Almost all the effort in thi…

DGX agent

Codex is overoptimised for large models: it ranks 2nd out of 10 for GLM 5.2 but drops to 9th place for Gemma-4! Almost all the effort in this field goes into tuning the weights. We wanted to know how

model-releasesclem-delangue--x
7 Aug 2026
Model Releases

Conditional Cognitive Biases in LLMs: How Biased User Turns Modulate In-Context Reasoning

DGX agent

arXiv:2608.05166v1 Announce Type: new Abstract: We present an evaluation of cognitive bias expression in state-of-the-art instruction-tuned LLMs under realistic multi-turn interaction settings. Our wo

model-releasesarxiv-cs-cl
7 Aug 2026
Model Releases

Continual Learning in Transition

DGX agent

arXiv:2608.06216v1 Announce Type: cross Abstract: Classical continual learning (CL) has primarily focused on enabling models to update and retain knowledge through parameter-centric mechanisms, e.g.,

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

ConWriter: Transition-Constrained Stateful Long-Form Story Generation with Lightweight Neuro-Symbolic Consistency Control

DGX agent

arXiv:2608.05169v1 Announce Type: new Abstract: Long-form story generation requires models to preserve narrative consistency across extended contexts, yet existing prompting-based methods often accumu

model-releasesarxiv-cs-cl
7 Aug 2026
Model Releases

CREBench: Evaluating Large Language Models in Cryptographic Binary Reverse Engineering

DGX agent

arXiv:2604.03750v2 Announce Type: replace-cross Abstract: Reverse engineering (RE) is central to software security, particularly for cryptographic programs that handle sensitive data and are highly pr

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

CRINN: Contrastive Reinforcement Learning for Approximate Nearest Neighbor Search

DGX agent

arXiv:2508.02091v4 Announce Type: replace-cross Abstract: Approximate nearest-neighbor search (ANNS) algorithms have become increasingly critical for recent AI applications, particularly in retrieval-

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

Cross-Architecture Steering Transfer in Language Models: A Systematic Empirical Study

DGX agent

arXiv:2608.05164v1 Announce Type: new Abstract: Independently trained large language models may develop shared internal representations of semantic concepts despite architectural differences -- but wh

model-releasesarxiv-cs-cl
7 Aug 2026
Model Releases

D-CLOT: Double Closed Loop Optimal Transport for Unsupervised Action Segmentation

DGX agent

arXiv:2608.05877v1 Announce Type: cross Abstract: Optimal transport (OT) has emerged as an effective framework for unsupervised action segmentation. Yet, in existing OT-based methods, the latent actio

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

DASH: Decoupled Adaptive Surrogate - Acquisition Harness for Automated Bayesian Optimization

DGX agent

arXiv:2608.00641v2 Announce Type: replace Abstract: Bayesian optimization (BO) relies on a surrogate model and an acquisition function, yet the most suitable choices vary across tasks and optimization

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

DASH: Divergence-Adaptive Supervision Horizons for On-Policy Self-Distillation of Reasoning Models

DGX agent

arXiv:2608.06243v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) improves the reasoning capabilities of large language models using automatically verifiable outcom

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

Decoupling Perception from Description: Computation-Grounded Representation Alignment between Multivariate Time Series and Language

DGX agent

arXiv:2608.05238v1 Announce Type: new Abstract: Training multimodal models to align time series with language runs into a self-supervision trap. The usual recipe asks an LLM to read a series and write

model-releasesarxiv-cs-lg
7 Aug 2026
Model Releases

Deep Generalised Mixed Models: a Novel Neural Network Structure for Analysing Hierarchical Data

DGX agent

arXiv:2608.05930v1 Announce Type: cross Abstract: The experience sampling method (ESM) is a longitudinal research design where participants report their thoughts, emotional states and behaviours multi

model-releasesarxiv-cs-lg
7 Aug 2026
Model Releases

DeepSeek-V4-Flash-0731 is now fully rolled out as the new default for deepseek-v4-flash on Ollama's cloud. This model combines speed, effici…

DGX agent

DeepSeek-V4-Flash-0731 is now fully rolled out as the new default for deepseek-v4-flash on Ollama's cloud. This model combines speed, efficiency, and frontier-level performance. Fast: 120+ output tps

model-releasesollama--x
7 Aug 2026
Model Releases

Domain-Grounded Candidate Selection for Agentic Image Editing: A Shadow Removal Case

DGX agent

arXiv:2608.06075v1 Announce Type: cross Abstract: Commercial vision-language models are reshaping computer vision, with visual priors broad enough to rival task-specific systems. This raises a natural

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

DREAM: LLM-based Dynamic Role-playing via Event-Aware Memory Graph

DGX agent

arXiv:2608.05170v1 Announce Type: cross Abstract: Role-playing agents (RPAs) have emerged as a key application of large language models, enabling immersive and high-fidelity character simulation. Accu

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

Dual-space posterior sampling for Bayesian inference in constrained inverse problems

DGX agent

arXiv:2603.00393v2 Announce Type: replace-cross Abstract: Inverse problems constrained by partial differential equations are often ill-conditioned due to noisy, incomplete data or inherent non-uniquen

model-releasesarxiv-cs-lg
7 Aug 2026
Model Releases

Dynamic Graph Prompting via Topology-Routed Mixed-Curvature Experts

DGX agent

arXiv:2608.06031v1 Announce Type: new Abstract: Dynamic graph prompting freezes a pre-trained temporal backbone and adapts it to label-scarce downstream tasks using lightweight prompts. However, exist

model-releasesarxiv-cs-lg
7 Aug 2026
Model Releases

DynaPix: Can Vision-Language Models Identify the Exact Future?

DGX agent

arXiv:2608.05505v1 Announce Type: new Abstract: Acting in a physical scene requires knowing its real later state, not a plausible one. Current evaluations often accept words or a realistic-looking ima

model-releasesarxiv-cs-cv
7 Aug 2026
Model Releases

ECHO: A Locally-Deployable Agentic Health Assistant with Temporal Memory, Safety Guardrails, and Speech Assessment

DGX agent

arXiv:2608.06110v1 Announce Type: new Abstract: This paper presents ECHO (Enhanced Care & Health Observer), a locally-deployable conversational health assistant for long-term chronic care management.

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

Echo Dot 2 can run 28M LLM at decent speed

DGX agent

Code and instructions available here: https://github.com/albertoZurini/echo-dot-2-playground Hello there! After a few days of experimenting I was able to get a completely local voice pipeline running

model-releasesr-localllama
7 Aug 2026
Model Releases

EcoAgent-Bench: Evaluating Economic Decision-Making in Budget-Constrained LLM Agents

DGX agent

arXiv:2608.05519v1 Announce Type: new Abstract: Agent benchmarks usually measure task completion and treat resource use as an auxiliary statistic. In deployment, however, the choice among a local look

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

Effective pruning of task-trained recurrent neural networks using noisy fluctuations and connection rescaling

DGX agent

arXiv:2608.05464v1 Announce Type: cross Abstract: The pruning of network connections is key to brain function but, despite its importance, there exist few biologically-plausible pruning rules with dem

model-releasesarxiv-cs-lg
7 Aug 2026
Model Releases

Energy-Guided Flow Matching

DGX agent

arXiv:2608.05811v1 Announce Type: new Abstract: Pixel-space generative models bypass lossy latent compression, yet necessitate joint learning of global structure and fine-grained details in a high-dim

model-releasesarxiv-cs-cv
7 Aug 2026
Model Releases

Enhancing Anomaly Resilience in Research Networks: A Large-Scale Forecasting Benchmark for Dynamic Security Baselining

DGX agent

arXiv:2608.05605v1 Announce Type: cross Abstract: Research and Education Networks (RENs) serve as critical infrastructure for scientific discovery, yet they face a unique security paradox: their norma

model-releasesarxiv-cs-lg
7 Aug 2026
Model Releases

Enhancing Social Intelligence in LLMs with Hierarchical Reasoning and Utterance-Level Goal Rewarding

DGX agent

arXiv:2608.05832v1 Announce Type: new Abstract: Large language models (LLMs) excel in structured tasks but struggle with dynamic social interactions, where success requires long-term goal coordination

model-releasesarxiv-cs-cl
7 Aug 2026
Model Releases

EpiBench: Can LLMs Understand Epitopes for Antibody Drug Discovery?

DGX agent

arXiv:2608.06022v1 Announce Type: new Abstract: Epitopes determine where antibodies bind antigens and shape downstream therapeutic properties such as functional blockade and escape resistance, making

model-releasesarxiv-cs-cl
7 Aug 2026
Model Releases

Equation-Free Period-Aware Forecast-Error Contraction for Estimating Negative Largest Lyapunov Exponents from Short Trajectory Ensembles

DGX agent

arXiv:2608.05522v1 Announce Type: cross Abstract: Estimating positive largest Lyapunov exponents from data is comparatively natural because neighboring trajectories separate, whereas stable dynamics r

model-releasesarxiv-cs-lg
7 Aug 2026
Model Releases

EschaLabs/Qwen3.6-35B-A3B-Escha-W2 · Hugging Face

DGX agent

Hey peeps. I know you're tired of low quants giving hard to believe numbers. I'm quite skeptical too and from what I tried I'm often left with the impression that the claims fall short. So this model

model-releasesr-localllama
7 Aug 2026
Model Releases

Evaluating and Improving Pedagogical Fit in LLM-Based AI Tutors with the Pedagogical Suitability Index

DGX agent

arXiv:2608.05411v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used as AI tutors, but a correct answer is not always a pedagogically appropriate one. In classroom learni

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

Evaluating Investment Logic in Large Language Models: A Real-World Benchmark Towards Personalzied Financial Agents

DGX agent

arXiv:2608.06108v1 Announce Type: new Abstract: Investment competence is inherently personalized: the same market evidence can justify different actions for investors with different goals, horizons, p

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

Everyone should pay attention to the training timelines in this video. OpenAI shares they started training a new internal model May 7. That …

DGX agent

Everyone should pay attention to the training timelines in this video. OpenAI shares they started training a new internal model May 7. That is more than two months before they released GPT-5.6 publicl

model-releasesallie-k--miller--x
7 Aug 2026
Model Releases

Evidence Lock Before Commitment: A Frozen Interface Degrades LLM-as-Judge Evaluation

DGX agent

arXiv:2608.05353v1 Announce Type: new Abstract: LLM judges are often asked to extract criteria and evidence before choosing between candidate answers. This workflow assumes that the intermediate recor

model-releasesarxiv-cs-cl
7 Aug 2026
Model Releases

Evidential Rule Learning for Interpretable Classification with Abstention

DGX agent

arXiv:2608.05859v1 Announce Type: cross Abstract: Interpretable classification often requires more than accurate predictions for real-life deployment: models should be transparent about the evidence b

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

EvReflection: Event-Driven Micro-Dynamics for Reflection Removal

DGX agent

arXiv:2608.06184v1 Announce Type: new Abstract: Despite remarkable progress in reflection removal, current methods primarily exploit static image priors from a single frame and still suffer from sever

model-releasesarxiv-cs-cv
7 Aug 2026
Model Releases

FinEvo-Bench: A Longitudinal Benchmark for Self-Evolving Agents in Professional Financial Workflows

DGX agent

arXiv:2608.06144v1 Announce Type: new Abstract: Most agent benchmarks evaluate tasks independently and cannot measure whether experience from one task helps with later tasks. Existing self-evolution b

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

FormBharo: Designing and Evaluating a Voice Agent for Conversational Form Filling in Rural India

DGX agent

arXiv:2608.06027v1 Announce Type: cross Abstract: In India, almost every social benefit starts with a form, yet the people who need these benefits most are often unable to read or write. Reaching them

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

From Sports to Safety: Benchmarking Proactive Risk Inference in MLLMs

DGX agent

arXiv:2608.05560v1 Announce Type: cross Abstract: Timely anticipation of physical hazards is essential for real-world safety, yet existing MLLM evaluations focus on harmful content or general risks, l

model-releasesarxiv-cs-cl
7 Aug 2026
Model Releases

GAUGE: A Measurement-Grounded Benchmark for Physical Fidelity in Simulation Engines and Video World Models

DGX agent

arXiv:2608.05948v1 Announce Type: new Abstract: Physics engines facilitate large-scale training and evaluation for embodied intelligence, while generative video world models are emerging as implicit s

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

Gemma 4 QAT could be improved further by Google aligning the QAT model to modern q4_k instead of q4_0

DGX agent

Hello, For the past few days I have been benchmarking Gemma 4 26b QAT UD Q4_K_XL extensively versus Bartowski's Q4_K_L. While QAT is certainly very effective and reducing memory consumption versus the

model-releasesr-localllama
7 Aug 2026
Model Releases

good grok

DGX agent

good grok Best match for this hierarchical hands-free setup: - Runtime: ActiveGraph (event-sourced log as source of truth) or LangGraph for supervisor/manager graphs - Roles as skills: Claude Agent SD

model-releasesyohei-nakajima--x
7 Aug 2026
Model Releases

Got job as Director of AI and Systems development self-taught

DGX agent

Hey everyone, I just wanted to share my journey here for some motivation. Three years ago, I saw the sudden spike in AI and realized it was the future of tech. My goal at the time was to be an indie g

model-releasesr-localllama
7 Aug 2026
Model Releases

GrandCode: Achieving Grandmaster Level in Competitive Programming via Agentic Reinforcement Learning

DGX agent

arXiv:2604.02721v3 Announce Type: replace Abstract: Competitive programming remains one of the last few human strongholds in coding against AI. The best AI system to date still underperforms the best

model-releasesarxiv-cs-ai
7 Aug 2026
← Previous
1…2627282930…465
Next →