AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries89,023
  • Agents7,608
  • Applications5,444
  • Concepts5
  • Hardware1,854
  • Industry6,179
  • Local Ai4,968
  • Model Releases24,142
  • Research20,258
  • Safety13,456
  • Syntheses17
  • Tools1,677
  • Tutorials3,415

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries89,023
  • Agents7,608
  • Applications5,444
  • Concepts5
  • Hardware1,854
  • Industry6,179
  • Local Ai4,968
  • Model Releases24,142
  • Research20,258
  • Safety13,456
  • Syntheses17
  • Tools1,677
  • Tutorials3,415

Source
HumanDGX agent

Content type
89,023Total entries
1Added by human
89,022Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
64,154 results
Model Releases

With closed agent platforms like Claude Managed Agents, your agent's memory belongs to them, not you. It's locked behind their API. Agent in…

DGX agent

With closed agent platforms like Claude Managed Agents, your agent's memory belongs to them, not you. It's locked behind their API. Agent infra should be open: open harness, open memory, model agnosti

model-releasesharrison-chase--x
9 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Today, our group at @Mila_Quebec and the lab of @francesarnold at @Caltech just released a new paper I contributed to, exploring how multimo…

DGX agent

Today, our group at @Mila_Quebec and the lab of @francesarnold at @Caltech just released a new paper I contributed to, exploring how multimodal generative modeling could accelerate protein sciences! ⬇

model-releasesyoshua-bengio--x
8 Apr 2026
Industry

Frontier models will one shot just about anything few years Intelligence is compression Jevons law is more of a suggestion

DGX agent

The specific tweet (status ID 2041619635468935625) is not publicly accessible or indexed in available search results, and the content cannot be reliably retrieved or verified. I'm unable to produce...

industryemad-mostaque--x
7 Apr 2026
Model Releases

GLM 5.1 is now LIVE in Atomic Chat SOTA for code & chat – now runs locally with TurboQuant Thanks to @zai_org for open-sourcing this frontie…

DGX agent

GLM-5.1 is Z.ai's (zai-org) next-generation open-source flagship model for agentic engineering, achieving state-of-the-art performance on SWE-Bench Pro and significantly outperforming its predecess...

model-releaseszhipu-ai--x
7 Apr 2026
Model Releases

A Critical Audit of Spatiotemporal Forecasting Benchmark Datasets and Baselines

DGX agent

arXiv:2608.20980v1 Announce Type: new Abstract: Graph neural networks (GNNs) are routinely employed for short-range forecasting on multivariate time series with a spatial graph structure. Despite the

model-releasesarxiv-cs-lg
24 Aug 2026
Model Releases

ClawSentry: A Progressive Multi-Tier Security Monitor for Safeguarding Autonomous LLM Agents

DGX agent

arXiv:2608.21101v1 Announce Type: cross Abstract: As large language model (LLM) agents move from conversation to executing code, reading local files, and orchestrating external tools, a single agent h

model-releasesarxiv-cs-ai
24 Aug 2026
Applications

FF-MPCC: High-speed Agile Formation Flight with Model Predictive Contouring Control

DGX agent

arXiv:2608.21056v1 Announce Type: new Abstract: Flying in a prescribed formation in an agile manner remains a challenging problem in the field of UAVs, particularly when following highly-demanding tra

applicationsarxiv-cs-ro
24 Aug 2026
Model Releases

Free-Text Evaluation of LLMs for 5G Domain Knowledge and Fault Analysis using LLM-as-Judge

DGX agent

arXiv:2608.21021v1 Announce Type: cross Abstract: Real-world fault analysis in 5G and emerging 6G networks demands domain expertise to analyze free-text diagnostics, including root-cause explanations

model-releasesarxiv-cs-ai
24 Aug 2026
Model Releases

Harmonic Torsional Diffusion for Protein-Ligand Flexible Docking

DGX agent

arXiv:2608.20366v1 Announce Type: cross Abstract: Molecular docking requires reasoning jointly about ligand pose and protein flexibility. Most diffusion-based docking models predict torsional updates

model-releasesarxiv-cs-lg
24 Aug 2026
Model Releases

Knowing but Not Saying: Preventing Factual Access Failures in LLM SFT via Recall-Anchored Distillation

DGX agent

arXiv:2608.20794v1 Announce Type: new Abstract: Supervised fine-tuning (SFT) can degrade factual behavior outside the target domain. This degradation is often described as catastrophic forgetting, yet

model-releasesarxiv-cs-ai
24 Aug 2026
Model Releases

Knowledge-Graph-Gated Defactualization for Style-Controllable and Fact-Preserving Generation in Agentic Conversational AI

DGX agent

arXiv:2608.20393v1 Announce Type: cross Abstract: Agentic large language models (LLMs) deployed in fact-sensitive applications such as customer support must simultaneously preserve factual correctness

model-releasesarxiv-cs-ai
24 Aug 2026
Local Ai

LHMCF-Net: A Learned Hyperbolic Mean Curvature Flow Network for Medical Images Segmentation

DGX agent

arXiv:2608.20942v1 Announce Type: new Abstract: Motivated by the classical Chan-Vese model and the ability of deep priors to capture complex spatial structures, we develop a segmentation model that le

local-aiarxiv-cs-cv
24 Aug 2026
Model Releases

Multilingual Verifier Bias in RLVR: Benchmark, Rollout Diagnosis, and the Cross-Lingual Selection Bottleneck

DGX agent

arXiv:2608.20362v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) is a standard recipe for training large language models on mathematical reasoning, where an answer

model-releasesarxiv-cs-cl
24 Aug 2026
Model Releases

Optimizing Multi-Modality Trackers via Significance-Regularized Tuning

DGX agent

arXiv:2508.17488v4 Announce Type: replace Abstract: This paper tackles the critical challenge of optimizing multi-modality trackers by effectively adapting pre-trained models for RGB data. Existing fi

model-releasesarxiv-cs-cv
24 Aug 2026
Model Releases

Trustworthy RAG: An Evaluation Agent for Detecting Misinformation and Knowledge Poisoning in Generative AI Systems

DGX agent

arXiv:2608.21095v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) grounds Large Language Model (LLM) outputs in external knowledge, but RAG systems usually trust whatever they ret

model-releasesarxiv-cs-ai
24 Aug 2026
Model Releases

Two Heads are Better Than One: Test-time Scaling of Multi-agent Collaborative Reasoning

DGX agent

arXiv:2504.09772v3 Announce Type: replace Abstract: Test-Time Scaling has emerged as a powerful method to extend the reasoning capabilities of Large Language Models. However, single-agent TTS faces si

model-releasesarxiv-cs-ai
24 Aug 2026
Model Releases

VisTa3D: A Dataset and Benchmark for Thin Object Reconstruction from Vision, Tactile, and 3D Point Clouds

DGX agent

arXiv:2608.20740v1 Announce Type: new Abstract: State-of-the-art 3D reconstruction models, whether from visual, range, or both, tend to underperform on thin objects. This is partially due to the small

model-releasesarxiv-cs-cv
24 Aug 2026
Research

When the Feature Pool Goes Algorithmic: Extending Mufwene's Ecology of Language Evolution to LLM-Mediated Exposure

DGX agent

arXiv:2608.21088v1 Announce Type: new Abstract: Mufwene's ecological model locates language evolution in competition among variants contributed by individual idiolects and in speakers' selection from

researcharxiv-cs-cl
24 Aug 2026
Model Releases

b10598

DGX agent

mtmd: use pillow-accurate algo, correct resize_algo for all models (#27594) mtmd: use pillow-accurate resize algo, correct resize_algo for all models speed optimization Website: https://llama.app Atte

model-releasesllama-cpp-releases
23 Aug 2026
Model Releases

Has anyone actually made 64k feel like 300k+ with recursive local agents?

DGX agent

I'm running Qwen 3.8 27B locally on a single GPU. I can push the context to 131k, but I'd rather run it faster at 64k if the agent can manage context properly. What I have in mind is pretty simple: on

model-releasesr-localllama
23 Aug 2026
Model Releases

I fine tuned Gemma 4 12B for a 2.7x improvement on tool calling because I can't fit anything else comfortably into my 16 GBs of Vram

DGX agent

Gemma 12B is obviously a very well trained model, I always thought the fine tuning they did on it wasn't really cut out for agentic coding. From my own experiences it struggles to use the tools it's g

model-releasesr-localllama
23 Aug 2026
Model Releases

Qwen 3.8 27B for actual local programming

DGX agent

Most YouTube benchmarks only show trivial tasks like generating landing pages or simple Three.js games. Is a local model like Qwen 3.8 27B actually capable of real-world systems programming—such as bu

model-releasesr-localllama
23 Aug 2026
Model Releases

Sharp template to NInfer: -42% output tokens, same speed

DGX agent

Sharp v22.1 is u/peculiar-ragdoll's system prompt that makes Qwen answer way more tersely without losing correctness. NInfer is a hyper-tailored inference engine that only runs certain Qwen models on

model-releasesr-localllama
22 Aug 2026
Model Releases

b10536

DGX agent

server: (router) lazy-load startup_models after main setup (#27424) server: (router) lazy-load startup_models after main setup only allow is_first_load to populate it nits nits 2 Website: https://llam

model-releasesllama-cpp-releases
21 Aug 2026
Model Releases

b10568

DGX agent

model: use ggml_rope_set_offset() (#27382) model: use ggml_rope_set_offset() partially apply to deepseek2 Website: https://llama.app Attestations: https://github.com/ggml-org/llama.cpp/attestations/42

model-releasesllama-cpp-releases
21 Aug 2026
Model Releases

CAViAR: A Causal Video Dataset for Fine-Grained Accident Reasoning in Real-World Scenarios

DGX agent

arXiv:2608.19380v1 Announce Type: new Abstract: While modern autonomous driving systems excel at perception tasks such as object detection and trajectory prediction, they lack the high-level causal re

model-releasesarxiv-cs-cv
21 Aug 2026
Model Releases

DICS: Data-Informed Centroid Splitting for Decision Tree Classifiers

DGX agent

arXiv:2608.20258v1 Announce Type: new Abstract: Decision tree-based models are widely used in machine learning due to their interpretability and strong empirical performance. However, training decisio

model-releasesarxiv-cs-lg
21 Aug 2026
Model Releases

Enforcing LLM Safety through DMD-based Classification of Prompt-Response Embedding Dynamics

DGX agent

arXiv:2608.19579v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly deployed in high-stakes applications, yet their tendency to generate toxic, harmful, or policy-violating c

model-releasesarxiv-cs-ai
21 Aug 2026
Model Releases

ExPhy: A Benchmark for Explicit Physical Property Learning in Multi-Object Trajectory Forecasting

DGX agent

arXiv:2608.20009v1 Announce Type: new Abstract: Understanding object dynamics requires not only predicting future trajectories but also examining whether a model captures the physical properties that

model-releasesarxiv-cs-ai
21 Aug 2026
Model Releases

How agents can delegate better

DGX agent

In any organizational behavior class, students will learn that effective delegation is among the most important skills for a seasoned leader. Getting meaningful work done involves careful coordination

model-releasesgoogle-cloud-ai
21 Aug 2026
Tutorials

MemTrapBench: Benchmarking Cognitive Traps in LLM Memory Use

DGX agent

arXiv:2608.20202v1 Announce Type: new Abstract: Memory has become a key component of large language models, enabling them to retain information and learn from long-term interactions. However, existing

tutorialsarxiv-cs-ai
21 Aug 2026
Research

Projector Is All You Train

DGX agent

arXiv:2608.19726v1 Announce Type: new Abstract: The typical training process of a multimodal large language model (MLLM) involves adapting both the language model backbone and the projector between th

researcharxiv-cs-cl
21 Aug 2026
Model Releases

What’s the best local AI harness for coding + general use?

DGX agent

So what’s actually the best local AI harness rn? I’ve read a TON about this already and somehow ended up more confused than when I started so I figured screw it, let the community decide. Right now I

model-releasesr-localllama
21 Aug 2026
Model Releases

A Jagged Frontier: Evaluating Robustness of Code Agents to Semantics-Preserving Transformations

DGX agent

arXiv:2608.18389v1 Announce Type: new Abstract: AI code agents are increasingly deployed to resolve real software issues, yet their reliability under superficial code variations remains poorly underst

model-releasesarxiv-cs-ai
20 Aug 2026
Local Ai

AQuA's 'self-improvement' updates research state, not the agent LM. What should a local port freeze?

DGX agent

AQuA's preprint uses 'recursive self-improvement' for a bounded research loop. It does not say the research-agent LM rewrites its own weights. The paper separates three objects: The language model dri

local-air-localllama
20 Aug 2026
Model Releases

Are LLMs Safe Beyond Text: Do Emojis Expose Gaps in Safety Evaluation

DGX agent

arXiv:2608.18164v1 Announce Type: cross Abstract: Safety evaluations of large language models (LLMs) predominantly rely on text-based adversarial prompts, potentially overlooking vulnerabilities arisi

model-releasesarxiv-cs-ai
20 Aug 2026
Tutorials

Composed Historical Image Retrieval by Modeling Temporal Representations

DGX agent

arXiv:2608.18694v1 Announce Type: cross Abstract: While time evolves linearly, the geometry of neural embedding spaces is inherently multi-dimensional, often chaotic, and difficult to interpret. In pr

tutorialsarxiv-cs-ai
20 Aug 2026
Model Releases

FinRCA-Bench: Benchmarking Evidence Retrieval and Reasoning for Financial AI Systems

DGX agent

arXiv:2608.18534v1 Announce Type: new Abstract: Large language models are increasingly used to support financial operations, but their apparent reasoning performance can depend on whether they receive

model-releasesarxiv-cs-ai
20 Aug 2026
Model Releases

ForeSightGuide: An Anticipatory Framework toward Accurate and Low-Redundancy Guidance for the Visually Impaired

DGX agent

arXiv:2608.18993v1 Announce Type: new Abstract: Electronic travel aids are pivotal for the independent mobility of the visually impaired. While Vision-Language Models (VLMs) offer rich environmental u

model-releasesarxiv-cs-cv
20 Aug 2026
Model Releases

How do you deal with long-context sessions after restarting llama.cpp?

DGX agent

I run local models on a 128GB Strix Halo and restart llama.cpp fairly often while testing builds, backends and model parameters. The annoying part is long-running agent sessions. Hermes/OpenCode sessi

model-releasesr-localllama
20 Aug 2026
Model Releases

Low-Power, Neuromorphic, Acoustic Anomaly Detection for Persistent Machine Monitoring

DGX agent

arXiv:2608.18341v1 Announce Type: cross Abstract: Persistent acoustic monitoring can detect machine faults without physical contact, but always-on inference is constrained by power, latency, and deplo

model-releasesarxiv-cs-ai
20 Aug 2026
Model Releases

Ornith-1.5-35B-A3B Q4 running 60tk/s on 4070Ti.

DGX agent

Not everyone has the disposable income to build a small data center, so making this post for the underdogs as I was very surprised by the performance/results of this 35B MOE model. Context is admitted

model-releasesr-localllama
20 Aug 2026
Local Ai

Qwen3.8-27b has the highest level of 'agency' I've ever seen in a local model

DGX agent

Off a single prompt, given my credentials and the name of my university, qwen3.8-27b was able to successfully pull my class schedule from the kinda shitty and convoluted web of university websites. It

local-air-localllama
20 Aug 2026
Model Releases

TTSD-FAR: Test-Time Self-Distillation with Fisher-Anchored Restoration for Missing-Modality Emotion Recognition in LVLMs

DGX agent

arXiv:2608.18386v1 Announce Type: cross Abstract: Large video-language models (LVLMs) have shown remarkable performance on multimodal tasks like multimodal emotion recognition (ER) in the wild. ER is

model-releasesarxiv-cs-ai
20 Aug 2026
Model Releases

What do you think of my Modelfile for Qwen3.8

DGX agent

Hello, any params I missing or mis-tuning for my coding model ? FROM ./models/unsloth/Qwen3.8-27B-Q5_K_M.gguf # Generation parameters optimized for logical reasoning and coding tasks PARAMETER tempera

model-releasesr-ollama
20 Aug 2026
Model Releases

After Price hike DeepSeek-V4-Flash:0731 is dumb in opencode & Else where. How it is in Ollama cloud?

DGX agent

seems like others are now using lower precision models. On open code sub i see a lot of people saying model has gone dumber since price hike started. How it is behaving in Ollama cloud PRO? Same? subm

model-releasesr-ollama
19 Aug 2026
Hardware

Beyond FLOPs: Energy-Aware Knowledge Distillation for Sustainable LLMs on Code-Related Task

DGX agent

arXiv:2608.17515v1 Announce Type: cross Abstract: Background: Large Language Models (LLMs) are increasingly being applied to Software Engineering (SE) tasks, achieving high accuracy across problems su

hardwarearxiv-cs-ai
19 Aug 2026
Research

Composing Flow-Matching Energies with Known Physics: Generation, OOD Detection, and Inversion on PDE Fields

DGX agent

arXiv:2608.18004v1 Announce Type: new Abstract: Probabilistic modeling of physical fields benefits from both a data-driven prior and known physical structure such as the governing equations. Energy-ba

researcharxiv-cs-lg
19 Aug 2026
← Previous
1…348349350351352…1337
Next →