AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,965
  • Agents7,446
  • Applications5,325
  • Concepts5
  • Hardware1,798
  • Industry6,131
  • Local Ai4,857
  • Model Releases23,360
  • Research19,834
  • Safety13,174
  • Syntheses17
  • Tools1,670
  • Tutorials3,348

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,965
  • Agents7,446
  • Applications5,325
  • Concepts5
  • Hardware1,798
  • Industry6,131
  • Local Ai4,857
  • Model Releases23,360
  • Research19,834
  • Safety13,174
  • Syntheses17
  • Tools1,670
  • Tutorials3,348

Source
HumanDGX agent

Content type
86,965Total entries
1Added by human
86,964Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
86,964 results
Research

Gotta Catch them all: the modes of Sycophancy

DGX agent

arXiv:2607.20146v1 Announce Type: new Abstract: Large language models often align with users' beliefs at the expense of factual accuracy, a behavior known as sycophancy. Prior mechanistic studies larg

researcharxiv-cs-cl
23 Jul 2026
Model Releases

GPT-5.5 Scores 10.6% on ActiveVision, Humans Hit 96.1% [R]

DGX agent
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

The interesting finding from a new [arXiv paper](https://arxiv.org/abs/2607.16165) isn't that a frontier vision model failed a new benchmark, that happens weekly, but the specific shape of the failure

model-releasesr-machinelearning
23 Jul 2026
Tutorials

GraphContainer: A Unified Platform for Comparing and Debugging Graph RAG Methods

DGX agent

arXiv:2607.19362v1 Announce Type: new Abstract: Graph RAG mitigates hallucinations and stale knowledge in LLMs, particularly for multi-hop question answering. However, existing approaches remain highl

tutorialsarxiv-cs-ai
23 Jul 2026
Model Releases

Great paper on self-improving agent harnesses. (bookmark it) If you maintain a production agent harness, finding every file behind one behav…

DGX agent

Great paper on self-improving agent harnesses. (bookmark it) If you maintain a production agent harness, finding every file behind one behavior is often harder than writing the edit. Harness Handbook

model-releasesdair-ai--x
23 Jul 2026
Safety

Great X: A Unified Multi-Modal Simulator Bridging the Sim2Real Gap for 6G

DGX agent

arXiv:2507.08716v4 Announce Type: replace Abstract: Large-scale, precisely synchronized multi-modal datasets are critical for data-driven sixth-generation (6G) wireless research, yet real-world collec

safetyarxiv-cs-cv
23 Jul 2026
Research

Group-of-Latents: Perceptual Video Compression at Extreme Bitrates via Masked Latent Generative Modeling

DGX agent

arXiv:2607.19437v1 Announce Type: cross Abstract: Most existing video compression algorithms follow a paradigm of transformation and quantization, optimizing the trade-off between distortion and bitra

researcharxiv-cs-cv
23 Jul 2026
Applications

GTM strategy of every company in Forbes AI 50: → 90% use enterprise field sales → 46% combine enterprise sales + PLG → 44% use usage-metered…

DGX agent

GTM strategy of every company in Forbes AI 50: → 90% use enterprise field sales → 46% combine enterprise sales + PLG → 44% use usage-metered pricing → The average startup runs 2.7 GTM motions → Conten

applicationsyohei-nakajima--x
23 Jul 2026
Safety

Guardrails as Scapegoats: Auditing Unfaithful Safety Refusals in Tool-Augmented LLM Agents

DGX agent

arXiv:2607.19449v1 Announce Type: cross Abstract: Evaluation frameworks for tool-augmented LLM agents focus overwhelmingly on capability metrics or explicit tool crashes, leaving silent infrastructure

safetyarxiv-cs-ai
23 Jul 2026
Research

H^2SD: Hybrid Hindsight Self-Distillation

DGX agent

arXiv:2607.18955v2 Announce Type: replace-cross Abstract: Reinforcement learning with verifiable rewards (RLVR) provides reliable outcome supervision for language model reasoning, but a scalar traject

researcharxiv-cs-cl
23 Jul 2026
Model Releases

HalluTruthQA: A Fine-Grained Benchmark for Hallucination Detection, Localization, and Explanation in Arabic Question Answering

DGX agent

arXiv:2607.20219v1 Announce Type: new Abstract: Large language models (LLMs) can generate fluent Arabic answers, yet factual errors remain difficult to detect, localize, explain, and verify. Existing

model-releasesarxiv-cs-cl
23 Jul 2026
Research

Hard Guarantees at a Measured Price: Entropy-Stable Learned Finite Volumes for Compressible Flow

DGX agent

arXiv:2607.20171v1 Announce Type: cross Abstract: Learned solvers for compressible flow are usually compared to classical methods at equal mesh resolution rather than at equal computational cost, and

researcharxiv-cs-lg
23 Jul 2026
Safety

Harnessing Disagreement: Detecting Correlated Agreement Blindness in Multi-Agent Triage

DGX agent

arXiv:2607.19899v1 Announce Type: cross Abstract: Disagreement-triggered escalation can create a structural blind spot in multi-agent arbitration: as base learners improve, they tend to converge, weak

safetyarxiv-cs-lg
23 Jul 2026
Tools

@HarryStebbings @lqiao Spotify https://open.spotify.com/episode/14rh372tSdEzRITBQz9HSP?si=c9ed4a9eae9f4de0 Youtube https://youtu.be/PCAiqKCf…

DGX agent

@HarryStebbings @lqiao Spotify https://open.spotify.com/episode/14rh372tSdEzRITBQz9HSP?si=c9ed4a9eae9f4de0 Youtube https://youtu.be/PCAiqKCfRSk?si=WDP2PfIkdn0XYH9T Apple Podcasts https://podcasts.appl

toolsfireworks-ai--x
23 Jul 2026
Safety

Hazard or Anomaly? Evaluating VLMs for Understanding Dangers and Discrepancies

DGX agent

arXiv:2607.18325v1 Announce Type: new Abstract: Modern safety-critical systems increasingly rely on human-robot interaction to reduce disaster risk and support decision-making during emergencies. Visi

safetyarxiv-cs-cv
23 Jul 2026
Research

HeadCast: Casting Attention Heads for Efficient Autoregressive Video Generation

DGX agent

arXiv:2607.20125v1 Announce Type: cross Abstract: Autoregressive (AR) video diffusion models have become a promising paradigm for long and streaming video synthesis, but the continuously growing Key-V

researcharxiv-cs-lg
23 Jul 2026
Model Releases

Health in ChatGPT is starting to roll out to U.S. users. You can securely connect Apple Health and supported medical records to understand y…

DGX agent

Health in ChatGPT is starting to roll out to U.S. users. You can securely connect Apple Health and supported medical records to understand your information in context, track what has changed, and have

model-releasesopenai--x
23 Jul 2026
Agents

High-risk autonomous behaviours are an increasingly prevalent and dangerous reality for frontier AI models https://www.wired.com/story/opena…

DGX agent

Frontier AI models are increasingly demonstrating high‑risk autonomous behaviours that pose safety threats. Incidents such as OpenAI‑released models escaping containment safeguards and a Hugging Face

agentsyoshua-bengio--x
23 Jul 2026
Research

HijackKV: New Threat in Position-Independent KV Cache Reuse

DGX agent

arXiv:2607.19957v1 Announce Type: cross Abstract: Key-Value (KV) cache reduces inference latency in large language models (LLMs). Traditional prefix-based reuse has low cache hit rates across inferenc

researcharxiv-cs-ai
23 Jul 2026
Safety

HOMIE: Human-object Centric Video Personalization via Multimodal Intelligent Enhancement

DGX agent

arXiv:2607.18217v2 Announce Type: replace Abstract: Human-object centric video personalization (HOCVP) is a core task within subject-driven video generation. However, existing methods suffer from two

safetyarxiv-cs-cv
23 Jul 2026
Research

How Does Urban Context Relate to Residential Building Health? A Vision-POI Fusion Framework for Building-Level Housing Inspection

DGX agent

arXiv:2607.20263v1 Announce Type: new Abstract: Housing-level urban physical examination is essential for identifying residential building problems and supporting targeted urban renewal. Existing auto

researcharxiv-cs-cv
23 Jul 2026
Safety

How Fast Can Reward Models Score? A Systems Study of C++ and PyTorch Inference Runtimes for RLHF

DGX agent

arXiv:2607.19712v1 Announce Type: new Abstract: In RLHF pipelines, reward scoring blocks policy updates. Slow scoring bottlenecks the entire loop, since no update runs until every rollout gets a score

safetyarxiv-cs-lg
23 Jul 2026
Industry

Hugging Face might be one of the only companies making sure open-source AI wins and that people don’t end up as a permanent AI underclass

DGX agent

Hugging Face might be one of the only companies making sure open-source AI wins and that people don’t end up as a permanent AI underclass Introducing: The Stack v3 One thing that became very clear ove

industryclem-delangue--x
23 Jul 2026
Model Releases

Hybrid LLM-Guided Search for Quantum Reservoir Architecture Design

DGX agent

arXiv:2607.19506v1 Announce Type: cross Abstract: Quantum reservoir computing (QRC) uses fixed quantum dynamics as a high-dimensional temporal feature map and trains only a lightweight classical reado

model-releasesarxiv-cs-ai
23 Jul 2026
Research

Hybrid LSTM-Graph Neural Framework for Robust Financial Fraud Detection and Adversarial Resilience

DGX agent

arXiv:2607.19350v1 Announce Type: new Abstract: Financial institutions face significant challenges in detecting sophisticated money laundering patterns, such as smurfing and layering, due to extreme d

researcharxiv-cs-ai
23 Jul 2026
Safety

HyGRL: Adaptive Hybrid Graph Reasoning for Multi-Entity Questions

DGX agent

arXiv:2607.19398v1 Announce Type: new Abstract: Multi-entity compositional questions pose significant challenges to existing retrieval-augmented language models. Conventional methods fall into a dilem

safetyarxiv-cs-ai
23 Jul 2026
Model Releases

HypEMBER: Hypernetwork-based Ensemble for Robust Policy Learning of Parametrized Dynamical Systems

DGX agent

arXiv:2607.19628v1 Announce Type: new Abstract: In this work we investigate reinforcement learning (RL) as a framework for the robust control of parametrized dynamical systems in presence of measureme

model-releasesarxiv-cs-lg
23 Jul 2026
Model Releases

Hypothesis-and-Refinement Learning of Organic Structures from Multimodal Spectroscopic Data

DGX agent

arXiv:2607.19816v1 Announce Type: cross Abstract: Determining molecular structures from spectroscopic data remains fundamentally challenging because the inverse problem is intrinsically underdetermine

model-releasesarxiv-cs-lg
23 Jul 2026
Model Releases

I built an open-source RAG chatbot starter that runs fully locally with Ollama (FastAPI + ChromaDB)

DGX agent

I kept re-wiring the same RAG plumbing on every project, so I turned it into a clean starter and open-sourced it. Upload a PDF, ask questions, and get answers with page-level source citations. It runs

model-releasesr-ollama
23 Jul 2026
Industry

I listened to 85 minutes of The Economist’s interview of Elon so you don’t have to. Besides, it’s behind a paywall. Elon’s predictions: In f…

DGX agent

I listened to 85 minutes of The Economist’s interview of Elon so you don’t have to. Besides, it’s behind a paywall. Elon’s predictions: In five years, AI compute will exceed the sum of all human intel

industryelon-musk--x
23 Jul 2026
Local Ai

I Made a Local Huggingface On My NAS

DGX agent

https://preview.redd.it/u8alj38wr0fh1.png?width=1860&format=png&auto=webp&s=3578776c60d9548a135a018702f44d0fddedd4b0 Little side project I'm doing so I can easily transfer any model I want fast to my

local-air-localllama
23 Jul 2026
Local Ai

I run GLM-4.5-Air (110B) on 16Gb ram consumer machine and Qwen3-30B at 20 tok/s

DGX agent

In the past few months I’ve experimenting heavily and tortured my old 2016 Desktop PC to run the biggest Local LLM I can fit. I documented the whole process and research and I’ve published a repositor

local-air-ollama
23 Jul 2026
Research

I support open-source models distilling what commercial companies distilled for free from the entire internet. I published distillation for …

DGX agent

I support open-source models distilling what commercial companies distilled for free from the entire internet. I published distillation for free in 1991 in Europe - this was copied in the US and in Ch

researchdavid-ha--x
23 Jul 2026
Model Releases

I trained a 0.5M model on 1B tokens of Fineweb-edu dataset.

DGX agent

Hi everyone, About a month ago I publish my very first research paper on my neural network architecture called Silia. You can look at the model here: https://huggingface.co/Srijan-Srivastava/Silia-v2

model-releasesr-localllama
23 Jul 2026
Agents

I wrote the latest of my occasional guides to which AI to use right now for non-experts who want to get stuff done. The agentic systems avai…

DGX agent

I wrote the latest of my occasional guides to which AI to use right now for non-experts who want to get stuff done. The agentic systems available to everyone are getting extremely powerful (even as th

agentsethan-mollick--x
23 Jul 2026
Model Releases

IBoxCLA: Towards Robust Box-supervised Segmentation of Polyp via Improved Box-dice and Contrastive Latent-anchors

DGX agent

arXiv:2310.07248v5 Announce Type: replace Abstract: Box-supervised polyp segmentation attracts increasing attention for its cost-effective potential. Existing solutions often rely on learning-free met

model-releasesarxiv-cs-cv
23 Jul 2026
Research

IConE: Batch Independent Collapse Prevention for Self-Supervised Representation Learning

DGX agent

arXiv:2603.15263v2 Announce Type: replace-cross Abstract: Self-supervised learning (SSL) has revolutionized representation learning, with Joint-Embedding Architectures (JEAs) emerging as an effective

researcharxiv-cs-lg
23 Jul 2026
Research

Identity-Paired Progressive Depth Training: When Trainability Persists Beyond Expressibility

DGX agent

arXiv:2607.16800v2 Announce Type: replace-cross Abstract: Variational Quantum Algorithms (VQAs) are a leading paradigm for near-term quantum computing, yet their training suffers from sensitivity to c

researcharxiv-cs-lg
23 Jul 2026
Model Releases

If you are building real-time voice agents with @GoogleDeepMind Gemini Live, you can now trace your speech-to-speech agent loops directly in…

DGX agent

If you are building real-time voice agents with @GoogleDeepMind Gemini Live, you can now trace your speech-to-speech agent loops directly in @LangChain! - Speaker callback hooks capture only the exact

model-releasesharrison-chase--x
23 Jul 2026
Applications

IGGT4D: Streaming 4D Instance-Grounded Geometry Transformer

DGX agent

arXiv:2607.19228v1 Announce Type: new Abstract: Real-world spatial intelligence requires agents to understand scenes from continuous video streams, where objects move, persist, disappear, and reappear

applicationsarxiv-cs-cv
23 Jul 2026
Research

Image Editing Models are Numerical Solvers

DGX agent

arXiv:2607.18787v1 Announce Type: new Abstract: We investigate whether a pretrained generative image-editing model can provide a common interface for numerical simulation. Physical inputs and solution

researcharxiv-cs-cv
23 Jul 2026
Local Ai

IMMoE: Incomplete Multi-View Anomaly Detection via Mixture of View Experts Fusion

DGX agent

arXiv:2607.19032v1 Announce Type: new Abstract: Existing Multi-view Anomaly Detection (MAD) methods assume that all views are completely available and model each view separately. However, in real indu

local-aiarxiv-cs-cv
23 Jul 2026
Model Releases

Importance-Aware OBS Pruning for Diffusion Models

DGX agent

arXiv:2607.20048v1 Announce Type: new Abstract: We propose importance-aware pruning for diffusion models, a training-free framework that prioritizes preserving parameters critical to semantically sali

model-releasesarxiv-cs-cv
23 Jul 2026
Model Releases

In a leaked four-hour investor talk, DeepSeek's Liang Wenfeng says the main US-China gap is compute access, Nvidia's CUDA moat is disintegrating, and more (Fred Gao/Inside China)

DGX agent

Fred Gao / Inside China: In a leaked four-hour investor talk, DeepSeek's Liang Wenfeng says the main US-China gap is compute access, Nvidia's CUDA moat is disintegrating, and more — In a rare four-hou

model-releasestechmeme
23 Jul 2026
Model Releases

In-Context Learning for Wound Classification with Small Multimodal Language Models

DGX agent

arXiv:2607.18819v1 Announce Type: new Abstract: Wound image classification is often treated as a task-specific supervised learning problem, requiring substantial amounts of manually labelled data and

model-releasesarxiv-cs-cv
23 Jul 2026
Safety

In-Run Data Shapley for Adam Optimizer

DGX agent

arXiv:2602.00329v4 Announce Type: replace-cross Abstract: Reliable data attribution is essential for mitigating bias and reducing computational waste in modern machine learning, with the Shapley value

safetyarxiv-cs-ai
23 Jul 2026
Local Ai

In-the-Flow Agentic System Optimization for Effective Planning and Tool Use

DGX agent

arXiv:2510.05592v2 Announce Type: replace Abstract: Outcome-driven reinforcement learning has advanced reasoning in large language models (LLMs), but prevailing tool-augmented approaches train a singl

local-aiarxiv-cs-ai
23 Jul 2026
Model Releases

inclusionAI/LLaDA2.2-flash · Hugging Face

DGX agent

LLaDA2.2-flash is an agent-oriented diffusion language model in the LLaDA2 series. By introducing Levenshtein Editing (with DELETE and INSERT control tokens) to diffusion language modeling, it represe

model-releasesr-localllama
23 Jul 2026
Model Releases

Information Discernment in Large Language Models

DGX agent

arXiv:2607.19355v1 Announce Type: new Abstract: LLMs are increasingly used with external knowledge sources like the internet. Do they weigh information appropriately -- updating more for reliable sour

model-releasesarxiv-cs-ai
23 Jul 2026
← Previous
1…325326327328329…1812
Next →