AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,510
  • Agents7,405
  • Applications5,305
  • Concepts5
  • Hardware1,789
  • Industry6,120
  • Local Ai4,835
  • Model Releases23,219
  • Research19,716
  • Safety13,102
  • Syntheses17
  • Tools1,670
  • Tutorials3,327

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Categories
  • All entries86,510
  • Agents7,405
  • Applications5,305
  • Concepts5
  • Hardware1,789
  • Industry6,120
  • Local Ai4,835
  • Model Releases23,219
  • Research19,716
  • Safety13,102
  • Syntheses17
  • Tools1,670
  • Tutorials3,327

Source
HumanDGX agent

86,510Total entries
1Added by human
86,509Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
86,510 results
9 Jul 2026

u guys were clowning on @greptile but turns out they were just the inspo for @openai to go so, so much harder https://x.com/JangLawrenceK/st…

Model ReleasesDGX agent

u guys were clowning on @greptile but turns out they were just the inspo for @openai to go so, so much harder https://x.com/JangLawrenceK/status/2075204015890325703 guys I just cancelled my Claude pla

UASPL: Uncertainty-Aware Self-Paced Learning with Evidential Neural Networks

ResearchDGX agent

arXiv:2607.06638v1 Announce Type: new Abstract: Self-paced learning (SPL) is an effective learning paradigm that simulates the human learning process by progressing from easy to difficult samples base

Understanding Interpretation Difficulty in Harmful Online Communication: Insights from Cybercrime Communities

Local AiDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub

arXiv:2607.07277v1 Announce Type: new Abstract: Harmful online communication often contains slang, coded terms, abbreviations, and community-specific expressions, which make messages difficult to inte

Understanding Two-Layer Neural Networks with Smooth Activation Functions

ResearchDGX agent

arXiv:2507.14177v2 Announce Type: replace-cross Abstract: This paper aims to understand the training solution, which is obtained by the back-propagation algorithm, of two-layer neural networks whose h

Unmasking On-Policy Distillation: Where It Helps, Where It Hurts, and Why

SafetyDGX agent

On-policy distillation offers dense, per-token supervision for training reasoning models; however, it remains unclear under which conditions this signal is beneficial and under which it is detrimental

Unraveling Machine Behavior by Multi-Level Bias Analysis and Detection: Methodology and Application to Computer Vision

Model ReleasesDGX agent

arXiv:2607.07236v1 Announce Type: new Abstract: This study investigates the presence and propagation of bias within Neural Networks through a comprehensive multi-level analysis spanning the learned la

UP: Unbounded Positive Asymmetric Optimization for Breaking the Exploration-Stability Dilemma

SafetyDGX agent

arXiv:2607.06987v1 Announce Type: new Abstract: Reinforcement learning (RL) has become the standard paradigm for enhancing the complex reasoning capabilities of large language models (LLMs). To achiev

URS-Stereo: Uncertainty-Guided Residual Search for Real-Time Stereo Matching

Local AiDGX agent

arXiv:2607.06779v1 Announce Type: new Abstract: Real-time stereo matching is crucial for robotics, autonomous systems, and embedded vision applications, where both computational efficiency and dispari

US cybersecurity and data resilience company Rubrik plans to invest $500M+ in the UK over the next five years and establish its European headquarters in London (Paul Sandle/Reuters)

IndustryDGX agent

Paul Sandle / Reuters: US cybersecurity and data resilience company Rubrik plans to invest 500M+ in the UK over the next five years and establish its European headquarters in London — U.S. cybersecuri

Use Grok 4.5 to build full-stack apps with Convex

IndustryDGX agent

Use Grok 4.5 to build full-stack apps with Convex Wow Grok 4.5 is very impressive at @convex code. Almost perfect score with a very very low run cost! Nice work @SpaceXAI! Without the guidelines it is

Validate the Dream Before You Trust Its Verdict: Admissibility for World-Model Simulators

SafetyDGX agent

arXiv:2607.07196v1 Announce Type: cross Abstract: Across robotics, World Models (WMs) are increasingly used to evaluate action policies by simulating the consequences of actions in an imagined world,

Value of Information under Imprecise Probabilities: Decision-Rule-Specific Values and Fixed-Measure Envelopes on a Credal Set

ResearchDGX agent

arXiv:2607.06570v1 Announce Type: cross Abstract: Value-of-information (VOI) analysis is usually conducted under a single probability measure. However, in practice, the available evidence often pins t

VCDP: Variation-Conditioned Distributional Proxy Learning for Semi-Supervised Medical Image Segmentation

Model ReleasesDGX agent

arXiv:2607.07416v1 Announce Type: new Abstract: Semi-supervised 3D medical image segmentation reduces the need for dense voxel-level annotations by exploiting unlabeled volumes. Although existing meth

Very cool to see @netflix releasing video datasets and models on @huggingface. They have a strong video AI team and would be amazing to see …

IndustryDGX agent

Netflix has released video datasets and models on Hugging Face, leveraging their internal video AI expertise to contribute to the open-source community. The post, shared by Clem Delangue (CEO of Huggi

VFM-Loc: Training-Free Cross-View Geo-Localization via Aligning Discriminative Visual Hierarchies

SafetyDGX agent

arXiv:2603.13855v2 Announce Type: replace Abstract: Cross-View Geo-Localization (CVGL) in remote sensing aims to locate a drone-view query by matching it to geo-tagged satellite images. Although super

Video-Based Detection of squint and cataract for accessibility-aware adaptive web interface rendering

ResearchDGX agent

arXiv:2607.07099v1 Announce Type: new Abstract: Squint and cataract are major ocular disorders that majorly affect visual perception and interaction capability. This paper proposes a real-time video-b

Video2Reaction: Mapping Video to Audience Reaction Distribution in the Wild

Model ReleasesDGX agent

arXiv:2607.06875v1 Announce Type: new Abstract: Understanding and forecasting audience reactions to video content are crucial for improving content creation, recommendation systems, and media analysis

Vision Foundation Models in Radiology: A Scoping Review of Data, Methodology, Evaluation and Clinical Translation

SafetyDGX agent

arXiv:2607.07219v1 Announce Type: cross Abstract: Vision foundation models (VFMs) are increasingly being developed for radiological imaging, yet their definition, development and evaluation remain het

Vision Language Action (VLA) Models for Unmanned Aerial Robotics and Bimanual Manipulation: A Review

ResearchDGX agent

arXiv:2607.06706v1 Announce Type: cross Abstract: Vision Language Action (VLA) models unify visual perception, natural-language understanding, and action generation within a single foundation model, a

VOTE: Vision-Language-Action Optimization with Trajectory Ensemble Voting

ResearchDGX agent

arXiv:2507.05116v5 Announce Type: replace-cross Abstract: Recent large-scale Vision Language Action (VLA) models have shown superior performance in robotic manipulation tasks guided by natural languag

VultronRetriever: Open Visual Document Retrieval Models Built for Scale

ApplicationsDGX agent

Introducing VultronRetriever, a family of open visual document retrieval models built on Vultr. Achieve state-of-the-art retrieval performance with smaller indexes, higher throughput, and support for

Wally Funk waited 60 years to get to space, and no one ever earned it more. She trained with the Mercury 13 in 1961, out-tested the men, and…

IndustryDGX agent

Wally Funk waited 60 years to get to space, and no one ever earned it more. She trained with the Mercury 13 in 1961, out-tested the men, and was told no anyway. She never stopped flying — 19,600 hours

WAM-TTT: Steering World-Action Models by Watching Human Play at Test Time

ResearchDGX agent

arXiv:2607.06988v1 Announce Type: cross Abstract: Steering robot foundation models (RFMs) toward new task variants or user-preferred behaviors remains challenging, often requiring additional robot dem

Watch Ollama's @jmorgan and @peterfenton on @CNBC with @dee_bosa at 12pm PT / 3pm ET to discuss 'The Post-Frontier Era' and 'Open-Source AI'…

Model ReleasesDGX agent

Watch Ollama's @jmorgan and @peterfenton on @CNBC with @dee_bosa at 12pm PT / 3pm ET to discuss 'The Post-Frontier Era' and 'Open-Source AI's Breakout Moment' Benchmark’s @peterfenton says 90%+ of tok

We are nearing 1,000 signups for the first-ever global video hackathon taking place next weekend. We are also opening a new track: VIDEO APP…

IndustryDGX agent

We are nearing 1,000 signups for the first-ever global video hackathon taking place next weekend. We are also opening a new track: VIDEO APPS. What new apps can you make w/ gen video? Filmmakers, deve

We built a plugin that traces every Claude Code session straight into LangSmith. Three commands, one JSON block, and every message, tool cal…

Model ReleasesDGX agent

We built a plugin that traces every Claude Code session straight into LangSmith. Three commands, one JSON block, and every message, tool call, and subagent run shows up as an inspectable trace. Setup

We comprehensively benchmarked GPT-5.6 on document understanding. At a high-level there's no change between GPT-5.6 Sol and GPT-5.5 in terms…

Model ReleasesDGX agent

We comprehensively benchmarked GPT-5.6 on document understanding. At a high-level there's no change between GPT-5.6 Sol and GPT-5.5 in terms of performance over tables, text, charts, layout, and more.

We evaluated 30+ frontier embodied AI models. The result is clear: current generalist robot policies are still far from robust real-world ma…

ApplicationsDGX agent

A comprehensive evaluation of over 30 frontier embodied AI models reveals that current generalist robot policies lack the robustness needed for reliable real-world manipulation tasks. The analysis dem

we have heard enterprises on their concerns about AI costs, and 5.6 sol is a huge step forward for dollars-per-task, as are terra and luna

IndustryDGX agent

Sam Altman discusses enterprise concerns about AI costs, highlighting that a 5.6 SOL rate represents significant progress in improving cost-efficiency metrics for AI task execution. The post reference

🤗 we just opensourced our tiny beast on @huggingface >0.9B ASR model >Apache license 2.0 >128k long-context transcription >Up to ~90-min au…

HardwareDGX agent

🤗 we just opensourced our tiny beast on @huggingface >0.9B ASR model >Apache license 2.0 >128k long-context transcription >Up to ~90-min audio input >Speaker labels + timestamps in one generation >Mul

We measured cost per task against GLM 5.2 as the baseline. On WANDR, GLM 5.2 + advisor runs at 2.1x versus Opus at 6.1x, averaging roughly h…

ToolsDGX agent

Perplexity conducted a cost efficiency comparison of language models, using GLM 5.2 as the baseline metric for cost per task. Results showed GLM 5.2 with an advisor achieved 2.1x cost efficiency on th

We shipped 9k Reachy Minis. They generate 15k hours of conversation a month. On GPT-realtime, that would cost $45k per month. So we built ou…

IndustryDGX agent

We shipped 9k Reachy Minis. They generate 15k hours of conversation a month. On GPT-realtime, that would cost 45k per month. So we built our own: fully open, one line to migrate, 0.25/hour. Free on yo

We stand at a critical crossroads in the debate over AI governance in the United States, and it feels like we are inching closer to a very s…

TutorialsDGX agent

We stand at a critical crossroads in the debate over AI governance in the United States, and it feels like we are inching closer to a very serious battle over whether or not open source models will ev

Weight-Space Physics: Interpretable Hypernetworks for Lattice Quantum Field Theories

ResearchDGX agent

arXiv:2607.07127v1 Announce Type: cross Abstract: Lattice field theory is the workhorse of non-perturbative physics, used to simulate phenomena from the strong nuclear force to critical phenomena in m

We're hosting a meetup on agent memory and wikis July 28th at our SF office! Come hear @jacobtpl and myself talk about the frontier research…

AgentsDGX agent

We're hosting a meetup on agent memory and wikis July 28th at our SF office! Come hear @jacobtpl and myself talk about the frontier research going on in these areas right now. https://luma.com/mylwoab

We're releasing a research preview of a new orchestrator model in Perplexity Computer. The model is an adapted version of GLM 5.2, post-trai…

ToolsDGX agent

We're releasing a research preview of a new orchestrator model in Perplexity Computer. The model is an adapted version of GLM 5.2, post-trained for the Computer harness. It delivers near-frontier perf

what a good video

IndustryDGX agent

Sam Altman shared thoughts on the characteristics or qualities that define effective video content. The post likely discusses principles for creating compelling videos, whether from a technical, story

What Predicts Correctness in Text-to-SQL? A Selective-Prediction Study

Model ReleasesDGX agent

arXiv:2607.06799v1 Announce Type: cross Abstract: Evaluating uncertainty in AI-generated SQL queries requires estimating whether a query is correct, where correct means it executes to the same result

What's on My Network? Using Large Language Models to Identify Real-World IoT Devices at Scale

Model ReleasesDGX agent

arXiv:2510.13817v2 Announce Type: replace Abstract: The growth of IoT devices in shared environments has outpaced our ability to identify them, posing urgent risks to privacy, safety, and accountabili

When Agents Go Rogue: Activation-Based Detection of Malicious Behaviors in Multi-Agent Systems

Local AiDGX agent

arXiv:2607.06807v1 Announce Type: cross Abstract: While enabling effective collaboration on complex tasks, LLM-based Multi-Agent Systems (MAS) face critical security challenges due to vulnerabilities

When Agents Remember Too Much: Memory Poisoning Attacks on Large Language Model Agents

SafetyDGX agent

arXiv:2607.06595v1 Announce Type: cross Abstract: Personal AI agents powered by large language models can reason and act using available tools to access emails, manage calendars, and push code to remo

When Certificates Fail: A Unified Safety Framework for Embedded Neural Interface Models

SafetyDGX agent

arXiv:2607.06630v1 Announce Type: new Abstract: Formal robustness certificates for embedded neural-interface models can pass while task accuracy collapses: at perturbation budget e=0.25, EEGNet classi

When Distillation Breaks Motion Control: Restoring Generative Trajectories for Fast Video Generators

ResearchDGX agent

arXiv:2506.19348v2 Announce Type: replace Abstract: Training-free motion customization imposes motion patterns from reference videos onto video generators through test-time computation. Most existing

When Do Geometric Algebra Layers Beat Scalarization? A Controlled Study on SO(3)-Equivariant Vector Laws

Local AiDGX agent

arXiv:2607.06634v1 Announce Type: new Abstract: Compact networks built from Clifford algebra Cl(3,0) primitives are exactly SO(3)-equivariant and learn synthetic 3D vector laws from few samples. We as

When Does In-Context Search Help? A Sampling-Complexity Theory of Reflection-Driven Reasoning

Local AiDGX agent

arXiv:2607.06720v1 Announce Type: new Abstract: Training large language models (LLMs) with extended reasoning has enabled in-context search, in which models iteratively generate, critique, and revise

When Prompts Ignore Structure: Graph-Based Attribute Reasoning for Calibrated VLMs

ResearchDGX agent

arXiv:2607.07395v1 Announce Type: cross Abstract: Reliable confidence estimation remains a key limitation of test-time adaptation in vision-language models (VLMs), where prompt tuning improves zero-sh

Where Did the Variability Go? From Vibe Coding to Product Lines by Regeneration

ResearchDGX agent

arXiv:2606.19042v2 Announce Type: replace-cross Abstract: In vibe coding, an emerging AI-driven paradigm, an LLM generates an entire program from a natural language prompt, but what happens to the var

WHERE to Generate Matters: Budget-Aware Synthetic Augmentation for Label Skewed Federated Learning

SafetyDGX agent

arXiv:2607.06616v1 Announce Type: cross Abstract: Label skew in federated learning (FL) causes client drift and degrades global accuracy. Synthetic data augmentation can reduce this imbalance; however

Where to Intervene? Benchmarking Fairness-Aware Learning on Differentially Private Synthetic Tabular Data

Model ReleasesDGX agent

arXiv:2607.07471v1 Announce Type: cross Abstract: Machine learning models are increasingly deployed in high-stakes domains, raising concerns about both privacy and fairness. Differential Privacy (DP)

While the writing style of LLMs is still as recognizable as ever, a new trend is that humans have started organically writing like them, too…

TutorialsDGX agent

While the writing style of LLMs is still as recognizable as ever, a new trend is that humans have started organically writing like them, too (which makes sense: of course you would end up imitating th

'Whoever Jerry is, he was excellent.' That's a customer talking about an agent. @PodiumHQ's Walker Ward sat down with our COO @j_schottenste…

AgentsDGX agent

'Whoever Jerry is, he was excellent.' That's a customer talking about an agent. @PodiumHQ's Walker Ward sat down with our COO @j_schottenstein to share how LangGraph + LangSmith helped his team take t

Why Fake ? Unveiling the Semantic Vocabulary of Deepfake Detectors

Local AiDGX agent

arXiv:2607.07216v1 Announce Type: new Abstract: Deepfake (DF) technology poses a significant threat to information integrity, driving the need for robust detection methods. Most DF detectors only cons

Widest-Path Reachability Fields for Connectivity-Preserving Slender Structure Segmentation

ResearchDGX agent

arXiv:2607.07123v1 Announce Type: new Abstract: Segmenting slender curvilinear structures such as retinal vessels, cracks, and roads demands topological correctness, as even a single-pixel discontinui

WildCity: A Real-World City-Scale Testbed for Rendering, Simulation, and Spatial Intelligence

AgentsDGX agent

arXiv:2607.06838v1 Announce Type: new Abstract: Humans can navigate an unfamiliar city and gradually form a coherent spatial mental map spanning tens of square kilometers. Can AI build spatial represe

Wordle 1,845 5/6 ⬛⬛⬛⬛🟨 ⬛⬛🟨⬛🟨 🟨⬛🟨🟨⬛ ⬛🟨⬛🟨🟩 🟩🟩🟩🟩🟩

Model ReleasesDGX agent

This post documents a Wordle game result (puzzle #1,845) played by Anthropic, showing the progression of guesses through color-coded tile feedback until reaching the correct five-letter word solution

Yann LeCun claimed on social media [7] that my foundational 1990 paper on Neural World Models [1] 'was never accepted through peer review.' …

AgentsDGX agent

Yann LeCun claimed on social media [7] that my foundational 1990 paper on Neural World Models [1] 'was never accepted through peer review.' This is simply false. The core concepts from the tech report

Yes, AI has definitely made the cheating problem (including getting 'help,' not cheating) worse but the problem was bad already Doing homewo…

ApplicationsDGX agent

Yes, AI has definitely made the cheating problem (including getting 'help,' not cheating) worse but the problem was bad already Doing homework improved final test grades for 86% of college students st

Yesterday we launched SWE-1.7 built on the open-source Kimi K2.7. Concerns about Chinese base models are real: K2.7 completed 87% of tasks t…

AgentsDGX agent

Yesterday we launched SWE-1.7 built on the open-source Kimi K2.7. Concerns about Chinese base models are real: K2.7 completed 87% of tasks that other models refuse over human-rights concerns. We train

You can now deploy Lovable apps to Vercel

ApplicationsDGX agent

Vercel announced integration support allowing developers to deploy applications built with Lovable directly to the Vercel platform. This integration streamlines the deployment workflow by connecting L

You've heard of Infrastructure as Code- but agent evals can now ride your existing Terraform setup! I've been using the new LangSmith Terraf…

AgentsDGX agent

You've heard of Infrastructure as Code- but agent evals can now ride your existing Terraform setup! I've been using the new LangSmith Terraform provider to auto-provision online evals + monitoring ale

← Previous
1…310311312313314…1442
Next →