AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
All
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,272 results
Model Releases

Filing: DeepSeek has invested ~$20.8M in Unitree Robotics' Shanghai IPO and agreed to jointly develop AI models for humanoid machines (Eduardo Baptista/Reuters)

DGX agent

Eduardo Baptista / Reuters: Filing: DeepSeek has invested ~20.8M in Unitree Robotics' Shanghai IPO and agreed to jointly develop AI models for humanoid machines — Chinese artificial intelligence start

model-releasestechmeme
6 Aug 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Final optimization: from ~10 tok/s to ~15 tok/s on DeepSeek-V4-Flash-0731 at 128K ctx - 1 RTX 3090

DGX agent

J'ai consacré beaucoup de temps à l'optimisation de DeepSeek-V4-Flash-0731 GGUF sur une seule RTX 3090. Mon exigence absolue pour chaque configuration était la suivante : Le modèle doit rester utilisa

model-releasesr-localllama
6 Aug 2026
Model Releases

FinPerMA: A Theory-Informed, Event-Grounded Personalized-Memory Benchmark for LLM Agents

DGX agent

arXiv:2608.04095v1 Announce Type: new Abstract: Large language model (LLM) agents are increasingly used as personalized assistants in high-stakes domains such as financial advising, yet it remains unc

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables

DGX agent

arXiv:2608.04077v1 Announce Type: new Abstract: Evaluating financial AI agents requires criteria aligned with real professional work. Existing rubric methods typically derive criteria from task prompt

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

FinReportBench: Measuring and Improving Institution-Grade Financial Report Generation

DGX agent

arXiv:2608.04374v1 Announce Type: cross Abstract: Large language models can produce fluent financial analysis, but fluency alone does not establish whether a report is suitable for institutional deliv

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

FinRpt: Dataset, Evaluation System and LLM-based Multi-agent Framework for Equity Research Report Generation

DGX agent

arXiv:2511.07322v3 Announce Type: replace-cross Abstract: While LLMs have shown great success in financial tasks like stock prediction and question answering, their application in fully automating Equ

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

Foreseeing the Invisible: Amodal Reconstruction of Leaf Fossil Images

DGX agent

arXiv:2608.04423v1 Announce Type: new Abstract: Fossil leaves are rarely preserved whole -- sedimentary rock hides, breaks, and erodes the lamina, yet paleobotany depends on the complete shape and out

model-releasesarxiv-cs-cv
6 Aug 2026
Model Releases

Formal Analysis and Supply Chain Security for Agentic AI Skills

DGX agent

arXiv:2603.00195v2 Announce Type: replace-cross Abstract: 32 pages, 5 theorems with full proofs, 68 references, open-source tool: https://github.com/qualixar/skillfortify. v2: corrects the bibliograph

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

From Financial Sentiment Classification to Return Predictability: A QLoRA Benchmark of Large Language Models

DGX agent

arXiv:2608.04200v1 Announce Type: cross Abstract: Financial sentiment classifiers are commonly evaluated against human labels, but strong linguistic performance does not necessarily imply economically

model-releasesarxiv-cs-lg
6 Aug 2026
Model Releases

From Score Matrices to Football-Aware Match-State Simulation: An Auditable LLM Harness for Exact-Score Reranking

DGX agent

arXiv:2608.05030v1 Announce Type: new Abstract: Football score forecasting combines a strong statistical core with a difficult contextual edge. Dynamic Poisson-family models estimate team strength, ex

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

From Transparent Labware Segmentation to Collision Avoidance: A Real-Time Edge-Aware Perception Pipeline

DGX agent

arXiv:2608.04769v1 Announce Type: new Abstract: This paper presents an edge-aware instance segmentation framework that enables real-time robotic collision avoidance with transparent laboratory glasswa

model-releasesarxiv-cs-ro
6 Aug 2026
Model Releases

FUSEP: A Multi-Center Benchmark for Diverse Tasks in Early Pregnancy Fetal Ultrasound Screening

DGX agent

arXiv:2608.04766v1 Announce Type: cross Abstract: A large number of infants with congenital anomalies are born each year globally, especially in areas with underdeveloped medical resources. Currently,

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

Galaxy Phase-Space and Field-Level Cosmology: The Strength of Semi-Analytic Models

DGX agent

arXiv:2512.10222v2 Announce Type: replace-cross Abstract: Semi-analytic models are a widely used approach to simulate galaxy properties within a cosmological framework, relying on simplified yet physi

model-releasesarxiv-cs-lg
6 Aug 2026
Model Releases

GEB-Bench: Abstract Structures Told in Many Voices

DGX agent

arXiv:2608.04111v1 Announce Type: cross Abstract: Can a model look at a river delta and a lightning bolt and see that they share a structure? We introduce GEB-Bench, a benchmark whose unit is an abstr

model-releasesarxiv-cs-cl
6 Aug 2026
Model Releases

General Availability of Pinecone Nexus Proves Knowledge Drives Real Outcomes for Agentic AI

DGX agent

Pinecone announced the general availability of Pinecone Nexus, a knowledge engine that converts an enterprise’s proprietary data into governed, agent‑ready knowledge delivered through a single query c

model-releasespinecone
6 Aug 2026
Model Releases

Geometry-Informed Parameter-Efficient Fine-Tuning of Pre-trained Molecular GNNs for Blood-Brain Barrier Permeability Prediction

DGX agent

arXiv:2608.04257v1 Announce Type: new Abstract: Blood-brain barrier permeability (BBBP) prediction is a critical screening task in central nervous system drug discovery, where candidate molecules must

model-releasesarxiv-cs-lg
6 Aug 2026
Model Releases

GLM/Qwen Appreciation Post

DGX agent

https://preview.redd.it/o6ik6qboeohh1.png?width=1134&format=png&auto=webp&s=4016f26c50c1d93bd3d0c7e880e9b55a2d75310f I have been running Qwen3.6 27b for a little while (mostly coding tasks) and recent

model-releasesr-localllama
6 Aug 2026
Model Releases

Google is expanding its Gemini-powered Ask Maps feature in English to 130+ countries and adds agentic food ordering, hotel booking, and attraction searching (Jada Jones/ZDNET)

DGX agent

Jada Jones / ZDNET: Google is expanding its Gemini-powered Ask Maps feature in English to 130+ countries and adds agentic food ordering, hotel booking, and attraction searching — ZDNET's key takeaways

model-releasestechmeme
6 Aug 2026
Model Releases

@GoogleDeepMind Humanoid legs or wheeled rovers? Should robots be cracking eggs? Watch as the @GoogleDeepMind team behind Gemini Robotics 2 …

DGX agent

On July 30 Google AI published a video announcing **Gemini Robotics 2**, an intelligence layer developed by DeepMind that aims to bring autonomous robots closer to everyday human environments. The pos

model-releasesgoogle-ai--x
6 Aug 2026
Model Releases

Gradient Immunity: Null-Space Resistance to Malicious Fine-Tuning

DGX agent

arXiv:2608.05045v1 Announce Type: cross Abstract: Released aligned large language models remain vulnerable to malicious downstream finetuning. Existing defenses are largely designed for the fine-tunin

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

GRALS: GCN-Guided Redundancy-Aware Local Search for Minimum Vertex Cover

DGX agent

arXiv:2503.06396v2 Announce Type: replace Abstract: The minimum vertex cover (MVC) problem seeks to identify the smallest set of vertices that cover all edges in an undirected graph. As a fundamental

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

GUARD: Grounding Uncertainty and Ablation-Based Risk Detection for Diffusion-Based VLAs

DGX agent

arXiv:2608.04510v1 Announce Type: cross Abstract: Diffusion-based vision-language-action (VLA) policies can generate plausible actions even when their predictions are weakly grounded in the visual and

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

Hallucinations on the Board: Tool-Augmented Evaluation of LLM Chess Commentary

DGX agent

arXiv:2608.04240v1 Announce Type: cross Abstract: Superhuman game engines in domains like chess have made expert-level evaluations easily accessible, yet they communicate what is true without the natu

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

HelloWorld: Enabling Socially Interactive Characters in Video World Models

DGX agent

arXiv:2608.05070v1 Announce Type: new Abstract: Despite the remarkable recent progress of video world models, social interaction between users and the characters within these worlds remains unsupporte

model-releasesarxiv-cs-cv
6 Aug 2026
Model Releases

Helping Music Co-Creation Agents 'Listen' Well: Hierarchical Self-Supervised World Models for Understanding and Generation

DGX agent

arXiv:2608.04378v1 Announce Type: cross Abstract: Collaborative music agents need internal representations rich enough to support both understanding and generation, yet flexible enough for a workflow

model-releasesarxiv-cs-lg
6 Aug 2026
Model Releases

How come artificialanalysis.ai ranks Gemma4 above Qwen3.6 27b in SciCode

DGX agent

Just came across this coding benchmark: SciCode Artificialanalysis.ai reports a ranking which contradicts the feeling we've towards those models in real life coding. Is Gemma 4 really that good, or a

model-releasesr-localllama
6 Aug 2026
Model Releases

How is Deepseek v4 flash 0731 running on Ollama cloud?

DGX agent

I cancelled my pro plan ealier because I wanted to use new Deepseek v4 flash 0731 which was available on Openrouter through API only (not yet on ollama cloud at the time). The old Deepseek v4 flash/pr

model-releasesr-ollama
6 Aug 2026
Model Releases

How many people in this sub try to train their own AI from scratch on their systems just for fun and to test out techniques from research papers?

DGX agent

As for me, I own a system with an RTX 5090, Ryzen 9 9950X3D2, and 64 GB of DDR5. Every time I see research come out with a new way to train AI, I immediately think to try it on my system to see the re

model-releasesr-localllama
6 Aug 2026
Model Releases

HyPASE: Hyperbolic Geometry for Parameter-Efficient Speech Emotion Fine-Tuning Framework for Large Audio-Language Models

DGX agent

arXiv:2608.04351v1 Announce Type: cross Abstract: Large Audio-Language Models (LALMs) excel at general speech understanding; however, adapting them to fine-grained tasks like Speech Emotion Recognitio

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

I get that AI labs need to make money, but zero-warning price spikes are a nightmare for production builds

DGX agent

Seen a ton of posts today about the DeepSeek API price hike. Half the feed is doom-posting, the other half is explaining basic GPU economics. Honestly, I get the cost side. Sub-cent tokens were never

model-releasesr-localllama
6 Aug 2026
Model Releases

I ported vLLM's serving stack to C++20: 66 MiB binary, no Python at inference, output checked token-for-token against vLLM

DGX agent

I'm the author, so discount the enthusiasm accordingly. This is an unaffiliated community port, not endorsed by the vLLM project, which it uses to verify its correctness. What started it: I love vLLM,

model-releasesr-localllama
6 Aug 2026
Model Releases

In addition to the upgrade in intelligence with GPT-5.6 Luna, Free and Go users can now use the “Think” button for more reasoning on harder …

DGX agent

OpenAI has released GPT‑5.6 Sol, which powers both instant and deep‑reasoning modes for ChatGPT Plus and Pro customers, delivering fact‑centric responses. Free and Go tier users will receive unlimited

model-releasesopenai--x
6 Aug 2026
Model Releases

InsightEmb: Learning Action-Intent Embeddings for Agentic Insight Retrieval

DGX agent

arXiv:2608.04761v1 Announce Type: cross Abstract: Self-improving agents accumulate reusable insights from prior trajectories, making retrieval increasingly important for turning accumulated experience

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

IslamicTurathBench: A Multi-Task, Multi-Discipline Benchmark for Evaluating Large Language Models on the Islamic Scholarly Tradition (turath)

DGX agent

arXiv:2608.04703v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used for question answering, education, and research, including in religious and cultural domains where an

model-releasesarxiv-cs-cl
6 Aug 2026
Model Releases

Item Response Theory for AI Safety

DGX agent

arXiv:2608.05086v1 Announce Type: new Abstract: Language models differ in how safely they behave and these differences are measured by safety benchmarks. But aggregated benchmark scores are hard to tr

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

K-EXAONE 2.0 Technical Report

DGX agent

arXiv:2608.04505v1 Announce Type: new Abstract: This technical report presents K-EXAONE 2.0, an open-weight multilingual foundation model developed by LG AI Research as a step in our effort toward glo

model-releasesarxiv-cs-cl
6 Aug 2026
Model Releases

Kathleen Writes: Autoregressive Generation and Data Scaling Without Attention

DGX agent

arXiv:2608.04678v1 Announce Type: new Abstract: Papers 1-2 of the Kathleen series showed that a byte-level, attention-free architecture built from a wavetable encoder and multi-scale reverberant state

model-releasesarxiv-cs-cl
6 Aug 2026
Model Releases

Kitchen Robotic Manipulation utilizing Foundation Models

DGX agent

arXiv:2608.04042v1 Announce Type: new Abstract: Deploying robots in everyday human environments requires perception systems that are both robust and adaptable to diverse, dynamic conditions. In this w

model-releasesarxiv-cs-ro
6 Aug 2026
Model Releases

KV cache quantization benchmarks: 413 pairs tested on Qwen 3.6 27B, Gemma 4 31B. KLD with BeeLlama.cpp v0.4.0: KVarN 6-bit beats q8_0, precision tail 1024 dominates

DGX agent

Link to the article: KV Cache Quantization Benchmarks: KVarN, Precision Tail KLD benchmarks with BeeLlama.cpp v0.4.0, fork of llama.cpp with more KV cache quantization options. Models: Qwen 3.6 27B Q5

model-releasesr-localllama
6 Aug 2026
Model Releases

Label-Free Target-Domain Adaptation for Unconstrained Event-Image Feature Matching via Dual-Stage Distillation

DGX agent

arXiv:2607.10082v2 Announce Type: replace Abstract: Building pixel-level correspondence between event and image data is a fundamental task for multi-sensor systems. However, existing cross-modal match

model-releasesarxiv-cs-cv
6 Aug 2026
Model Releases

LaPrune: Controllable Differentiable Sparsity at Million Scale

DGX agent

arXiv:2608.04057v1 Announce Type: cross Abstract: Top-k selection determines which components of a sparse model remain active. Hard selection blocks gradients, while continuous relaxations often coupl

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

Large Language Models for Low-Resource Languages: A Conceptual Framework for an Electronic Explanatory Dictionary of the Tajik Language

DGX agent

arXiv:2608.04186v1 Announce Type: new Abstract: This paper presents a conceptual framework for developing an electronic explanatory dictionary of the Tajik language using large language models (LLMs).

model-releasesarxiv-cs-cl
6 Aug 2026
Model Releases

Large-Small Model Collaboration for Enhancing Edge-Deployed Small Models

DGX agent

arXiv:2503.10367v2 Announce Type: replace-cross Abstract: Edge devices host domain-specific small language models (SLMs) with limited resources, while private clouds offer larger LLMs. We propose G-Bo

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

Leak-Resistant Unlearning: A New Benchmark for Evaluating Multi-Hop Reasoning Consistency and Recovery Robustness

DGX agent

arXiv:2608.04519v1 Announce Type: new Abstract: Benchmarking machine unlearning methods is critical to understand whether sensitive knowledge is removed from large language models (LLMs) or not. Curre

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

LiNC: Lightweight Noise Correction via Per-Sample Trust and Gaussian Mixture Modeling

DGX agent

arXiv:2608.04147v1 Announce Type: cross Abstract: Label noise is common in medical imaging datasets due to factors such as inter-rater variability, annotation errors, and ambiguous cases. This can sev

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

Lindblad-Inspired Multi-Timescale Reservoir Computing with Separable Rotation and Dissipation

DGX agent

arXiv:2608.04028v1 Announce Type: cross Abstract: Echo-state networks enable efficient temporal learning by fixing the recurrent dynamics and training only a linear readout. However, conventional rese

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

LiveXiv -- A Multi-Modal Live Benchmark Based on Arxiv Papers Content

DGX agent

arXiv:2410.10783v4 Announce Type: replace Abstract: The large-scale training of multi-modal models on data scraped from the web has shown outstanding utility in infusing these models with the required

model-releasesarxiv-cs-cv
6 Aug 2026
Model Releases

LLM optimization integration for Amazon SageMaker Python SDK

DGX agent

The Amazon SageMaker Python SDK v3 now exposes generative AI inference recommendations in Amazon SageMaker AI directly in your notebook. Benchmark an endpoint, generate data-driven deployment recommen

model-releasesaws-ml-blog
6 Aug 2026
← Previous
1…3031323334…464
Next →