AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,919 results
Local Ai

Try Again, Don't Look Back: Blind Resampling Outperforms Self-Repair in Small Code Models

DGX agent

arXiv:2607.26117v1 Announce Type: cross Abstract: Self-repair - returning a failed program to the model together with its test output and asking for a correction - is a standard component of code agen

local-aiarxiv-cs-lg
30 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Local Ai

What actually happened to the whole Openclaw frenzy?

DGX agent

A while back you couldn't open reddit or youtube without sifting through tons of Openclaw content. And it wasn't just the internet that blew up, I remember seeing images from China where crowds would

local-air-localllama
30 Jul 2026
Research

A Cross-lingual Comparison of Human and Classification Model Entrainment Behavior in Code-switched Speech Settings

DGX agent

arXiv:2607.25202v1 Announce Type: new Abstract: Conversational entrainment is well-studied in monolingual and written contexts, but remains underexplored in spoken code-switching (CSW). We present a n

researcharxiv-cs-cl
29 Jul 2026
Model Releases

Automated Modernization of Machine Learning Engineering Notebooks for Reproducibility

DGX agent

arXiv:2602.07195v2 Announce Type: replace-cross Abstract: Interactive computational notebooks (e.g., Jupyter notebooks) are widely used in machine learning engineering (MLE) to program and share end-t

model-releasesarxiv-cs-lg
29 Jul 2026
Model Releases

Crystalis: Progressive Nucleation and Semantic Annealing for Coordinated Multi-View Visualization Generation

DGX agent

arXiv:2607.24766v1 Announce Type: new Abstract: Large language models (LLMs) can generate individual charts, but coordinated multi-view visualizations (CMVs), where views share data flows and cross-vi

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

DeepSeek V4 Flash isn't just for inference anymore. Fine-tune it on Fireworks with supervised fine-tuning, preference tuning, and combined p…

DGX agent

DeepSeek V4 Flash isn't just for inference anymore. Fine-tune it on Fireworks with supervised fine-tuning, preference tuning, and combined preference optimization from the managed UI. Reinforcement le

model-releasesfireworks-ai--x
29 Jul 2026
Safety

Do Models Fake Alignment Without Clear Consequences?

DGX agent

arXiv:2607.24758v1 Announce Type: new Abstract: Large language models are capable of recognizing evaluation contexts and altering their behavior to reflect evaluator expectations rather than typical d

safetyarxiv-cs-ai
29 Jul 2026
Safety

DSCD-Nav: Dual-Stance Cooperative Debate for Object Navigation

DGX agent

arXiv:2601.21409v3 Announce Type: replace Abstract: Adaptive navigation in unfamiliar indoor environments is crucial for household service robots. Despite advances in zero-shot perception and reasonin

safetyarxiv-cs-ro
29 Jul 2026
Local Ai

Everyone posts day-one impressions. What's still in your stack a month later?

DGX agent

Day one threads are the least useful thing we produce here and we produce a lot of them. Model drops, forty people run their favourite prompt, half say it's the best thing ever and half say benchmaxxe

local-air-localllama
29 Jul 2026
Research

FinAbstain: Uncertainty-Calibrated Multimodal RAG for Selective Financial Forecasting

DGX agent

arXiv:2607.24875v1 Announce Type: new Abstract: Large language models (LLMs) can synthesize financial narratives but may express high confidence when evidence is sparse, stale, or contradictory. This

researcharxiv-cs-lg
29 Jul 2026
Safety

NEXT: Reasoning-Driven Video Recommendation via a Vision-Language Model

DGX agent

arXiv:2607.24789v1 Announce Type: cross Abstract: We present NEXT (Next-interest EXploration Transformer), a reasoning-driven video recommendation framework that reasons over the video a user has just

safetyarxiv-cs-cv
29 Jul 2026
Safety

RoboHarness: Memory-Driven Orchestration of Heterogeneous Robot Policies for Long-Horizon Planning

DGX agent

arXiv:2607.18060v2 Announce Type: replace Abstract: Long-horizon robotic tasks require diverse capabilities that no single policy can reliably provide. Heterogeneous policies offer complementary stren

safetyarxiv-cs-ro
29 Jul 2026
Model Releases

Sense it with your eyes: Sensation Generation and Understanding for Advertisements

DGX agent

arXiv:2607.25314v1 Announce Type: new Abstract: Sensory advertising evokes human senses through visual cues, enabling audiences to mentally simulate experiences and increasing persuasive impact. Despi

model-releasesarxiv-cs-cv
29 Jul 2026
Model Releases

The median task consumed nearly 6x more tokens in Claude Code than in Kimi Code: - 61k in Kimi Code - 67k in Hermes - 340k in Claude Code At…

DGX agent

The median task consumed nearly 6x more tokens in Claude Code than in Kimi Code: - 61k in Kimi Code - 67k in Hermes - 340k in Claude Code At K3's 3/M input rate (input tokens make up roughly 95% of ag

model-releasesnous-research--x
29 Jul 2026
Model Releases

Thrilled to have @gabepereyra speak at our @sequoia event tmrw on OWN YOUR AI: how to build your own Lab as an application company. Also fea…

DGX agent

Thrilled to have @gabepereyra speak at our @sequoia event tmrw on OWN YOUR AI: how to build your own Lab as an application company. Also featuring @FireworksAI_HQ @mercor_ai @LangChain @trajectorylabs

model-releasessonya-huang--x
29 Jul 2026
Model Releases

Two of the people most responsible for scaling the transformer are now betting on a next act. @MillionInt ran the Reasoning 🍓 team at OpenA…

DGX agent

Two of the people most responsible for scaling the transformer are now betting on a next act. @MillionInt ran the Reasoning 🍓 team at OpenAI. @_arohan_ was a pre-training lead on Gemini after years at

model-releasessonya-huang--x
29 Jul 2026
Local Ai

Understand Kimi K3 from first principles: a recommended order for anyone trying to understand this beast

DGX agent

Everyone is talking about Kimi K3, but if you jump straight into the technical report, you’ll quickly realize it’s standing on years of research -- just like any breakthrough is! If you want to unders

local-air-localllama
29 Jul 2026
Model Releases

AI-generated Images Challenge Visual Trust in High-risk Scenarios

DGX agent

arXiv:2607.22745v1 Announce Type: cross Abstract: Rapid advances in image generation are eroding the evidentiary value of visual content in settings where authenticity can affect public safety and per

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

AIR-BENCH Live: An Evolving Safety Benchmark for Foundation Models

DGX agent

arXiv:2607.22671v1 Announce Type: new Abstract: Foundation-model safety benchmarks capture the AI risks of their time of publication: as models improve and governments pass new AI-safety legislation,

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

Alibaba has released Qwen Audio 3.0 Realtime, with the Plus variant debuting as the new #1 model on the Artificial Analysis Speech to Speech…

DGX agent

Alibaba has released Qwen Audio 3.0 Realtime, with the Plus variant debuting as the new #1 model on the Artificial Analysis Speech to Speech Index at 84.1%, ahead of GPT-Realtime-2.1 High at 79.1% Rel

model-releasesqwen--x
28 Jul 2026
Hardware

AMD calls its shot, but the real race is engineering velocity

DGX agent

At AMD’s Advancing AI event, Lisa Su called her shot – just like Babe Ruth. The question now is whether AMD has built the engineering machine to hit it. Last week, we argued that AMD’s next reinventio

hardwaresiliconangle
28 Jul 2026
Model Releases

CALMRec: Causally Aligned Language Memory for Long-Horizon Recommendation

DGX agent

arXiv:2607.23647v1 Announce Type: cross Abstract: Large language models (LLMs) can summarize heterogeneous user evidence in natural language, but current LLM recommenders often collapse enduring prefe

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

Chart Deception in Vision-Language Models: From Vulnerability to Mitigation

DGX agent

arXiv:2607.22600v1 Announce Type: new Abstract: Information visualizations are widely used to communicate patterns, trends, and outliers, yet deceptive design choices-such as truncated or inverted axe

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding

DGX agent

arXiv:2607.24743v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) hold immense potential to revolutionize clinical practice, yet deploying them in the medical domain is fundam

model-releasesarxiv-cs-ai
28 Jul 2026
Research

Co-Evolving Graph and Text Memory for Training-Free Multi-Hop Question Answering

DGX agent

arXiv:2607.23278v1 Announce Type: new Abstract: Multi-hop question answering requires coordinating relational and textual evidence across reasoning steps, a combination neither a text corpus nor a kno

researcharxiv-cs-cl
28 Jul 2026
Safety

Concept-based Visual Counterfactual Explanations with Diffusion Models

DGX agent

arXiv:2607.22544v1 Announce Type: new Abstract: Visual counterfactual explanations aim to answer 'what minimal change to this image would flip the model's prediction?', and are increasingly important

safetyarxiv-cs-ai
28 Jul 2026
Model Releases

Cortex: Compact Behavior Cloning for Quake with Frozen Visual Features

DGX agent

arXiv:2607.22739v1 Announce Type: cross Abstract: We study how far a deliberately simple behavioral-cloning policy can progress in a visually rich first-person game before adding reinforcement learnin

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

DeepLook: Deeper Thinking with Lookahead

DGX agent

arXiv:2607.22602v1 Announce Type: new Abstract: Inference-time scaling has emerged as a powerful paradigm for improving large language model reasoning, often delivering larger gains on difficult reaso

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

Do Current Retrievers Cover All the Evidence? A Controlled Study of Conjunctive Cross-Page Retrieval

DGX agent

arXiv:2607.24165v1 Announce Type: cross Abstract: Finding a long document relevant to a multi-part request is not the same as establishing that it contains every requested piece of evidence. We study

model-releasesarxiv-cs-cl
28 Jul 2026
Research

Do Visual Features Improve Other-Initiated Repair Detection? A Dyadic Multimodal Approach

DGX agent

arXiv:2607.23845v1 Announce Type: new Abstract: Other-initiated Self-repair, or in short Other-initiated Repair (OIR), is an essential mechanism in conversational interaction, whereby a recipient sign

researcharxiv-cs-ai
28 Jul 2026
Model Releases

DualityCert: Verifier-Gated Language-Model Repair of Broken Duality Claims in Quantum Field Theory

DGX agent

arXiv:2607.23614v1 Announce Type: cross Abstract: We present DualityCert, a symbolic verifier for candidate Seiberg-duality claims in four-dimensional N=1 quiver gauge theories. The verifier evaluates

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

EgoPlay: Event-Triggered Video Editing for Egocentric Streams

DGX agent

arXiv:2607.24560v1 Announce Type: cross Abstract: We introduce EgoPlay, an event-triggered video-to-video editor for egocentric streams, obtained by fine-tuning a pretrained V2V diffusion transformer

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

Embodied GPT-5.1: Evidence of a World Model?

DGX agent

arXiv:2607.23899v1 Announce Type: cross Abstract: This exploratory study examines whether a large multimodal language model, GPT-5.1, can serve as the high-level controller of a physical mobile robot

model-releasesarxiv-cs-ai
28 Jul 2026
Applications

Everything you need to know about Kimi K3 with @Kimi_Moonshot and @togethercompute team. If you're looking into evaling and building with Ki…

DGX agent

Everything you need to know about Kimi K3 with @Kimi_Moonshot and @togethercompute team. If you're looking into evaling and building with Kimi K3 I would not miss this one! Kimi K3 has everyone’s atte

applicationstogether-ai--x
28 Jul 2026
Safety

From Camera-Based Sensing to Reasoning: A Comprehensive Review Toward Proactive Vulnerable Road User Safety

DGX agent

arXiv:2510.03314v2 Announce Type: replace-cross Abstract: Ensuring the safety of vulnerable road users (VRUs), such as pedestrians and cyclists, remains a critical challenge, as conventional infrastru

safetyarxiv-cs-ai
28 Jul 2026
Research

From Execution to Capability: Scientific Experience Consolidation via Procedural Knowledge Synthesis

DGX agent

arXiv:2607.24459v1 Announce Type: new Abstract: Large language models increasingly solve scientific-computing tasks, but executable feedback from one problem rarely becomes durable capability on subse

researcharxiv-cs-ai
28 Jul 2026
Safety

From 'Help' to Helpful: A Hierarchical Assessment of LLMs in Mental e-Health Applications

DGX agent

arXiv:2602.18443v2 Announce Type: replace-cross Abstract: Psychosocial online counselling frequently encounters generic subject lines that impede efficient case prioritisation. This study evaluates el

safetyarxiv-cs-ai
28 Jul 2026
Safety

HALLELUAI: A Hallucination-Aware AI System for Ultra-Realistic Image-to-Video Generation at Scale

DGX agent

arXiv:2607.22959v1 Announce Type: cross Abstract: AI-generated video is increasingly used across marketing, product storytelling, and creative workflows, yet automated; high-precision quality control

safetyarxiv-cs-ai
28 Jul 2026
Model Releases

Hierarchical Group-Conditional Conformal Risk Control for Selective Prediction in Language Models

DGX agent

arXiv:2607.24562v1 Announce Type: new Abstract: Large language models serve heterogeneous populations structured by domain, topic difficulty, and linguistic style. Conformal risk control (CRC) gives r

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

How OpenAI hacked HuggingFace. What we know. Hugging Face proved that open platforms and open models can still win those battles when the al…

DGX agent

How OpenAI hacked HuggingFace. What we know. Hugging Face proved that open platforms and open models can still win those battles when the alternative is locked-down systems that refuse to assist their

model-releasesclem-delangue--x
28 Jul 2026
Model Releases

I built a tool to actually test which weights matter before quantizing, instead of guessing (Qwen3.6-27B, 3 builds: Bedrock/Tightrope/Gambit)

DGX agent

Most quantization works like this: pick a bit depth, apply it everywhere, maybe let imatrix take a rough guess at what matters, ship it. Most don't check which specific weight groups can take a hit an

model-releasesr-localllama
28 Jul 2026
Safety

In-Context Learning as Implicit Policy Gradient

DGX agent

arXiv:2607.23153v1 Announce Type: cross Abstract: Recent work has shown that large language models (LLMs) can iteratively improve their outputs by incorporating generated samples and their correspondi

safetyarxiv-cs-ai
28 Jul 2026
Applications

IndicTalk: A Large-Scale Persona-Based Multilingual Conversational Corpus for Indic Languages

DGX agent

arXiv:2607.23242v1 Announce Type: new Abstract: Large Language Models (LLMs) have transformed conversational AI, yet high-quality multilingual code-mixed dialogue resources remain scarce, particularly

applicationsarxiv-cs-cl
28 Jul 2026
Applications

Kimi K3 has everyone’s attention. On July 30, hear @Kimi_Moonshot's Feihu Tang explain the architecture and decisions behind it. He joins Ju…

DGX agent

Kimi K3 has everyone’s attention. On July 30, hear @Kimi_Moonshot's Feihu Tang explain the architecture and decisions behind it. He joins Jue Wang and Zain Hasan from Together AI for a deep dive into

applicationstogether-ai--x
28 Jul 2026
Model Releases

LEACL: LLM-Enhanced Automatic Curriculum Learning for Reinforcement Learning in Long-Horizon Manipulation Tasks

DGX agent

arXiv:2607.23515v1 Announce Type: new Abstract: Long-horizon manipulation tasks pose significant challenges for reinforcement learning due to sparse reward signals and long horizons. Automatic curricu

model-releasesarxiv-cs-ro
28 Jul 2026
Model Releases

LU-500: A Logo Benchmark for Concept Unlearning

DGX agent

arXiv:2607.24101v1 Announce Type: cross Abstract: Concept unlearning is increasingly used to limit the reproduction of protected or unsafe visual concepts in text-to-image models. Existing evaluations

model-releasesarxiv-cs-ai
28 Jul 2026
Safety

MEMENTO: Memory-Guided Memetic Code-as-Policy Evolution

DGX agent

arXiv:2607.22832v1 Announce Type: new Abstract: Long-horizon embodied tasks require policies that execute many dependent actions before task success can be observed. Representing policies as executabl

safetyarxiv-cs-lg
28 Jul 2026
Model Releases

microsoft/Mage-VL · Hugging Face - An Efficient Codec-Native Streaming Multimodal Foundation Model

DGX agent

Mage-VL is a codec-native, proactive-streaming multimodal foundation model for image and video understanding, whose visual encoder is trained entirely from scratch at a compact 4B scale. It targets a

model-releasesr-localllama
28 Jul 2026
← Previous
1…335336337338339…374
Next →