AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

91,060Total entries
1Added by human
91,059Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
91,060 results
3 Jun 2026

VidMsg: A Benchmark for Implicit Message Inference in Short Videos

Model ReleasesDGX agent

arXiv:2606.03635v1 Announce Type: cross Abstract: Understanding short online videos involves more than identifying visible objects and actions; video makers often include an underlying message or purp

VistaHop: Benchmarking Multi-hop Visual Reasoning for Visual DeepSearch

Model ReleasesDGX agent

arXiv:2606.03273v1 Announce Type: cross Abstract: Visual DeepSearch requires multimodal large reasoning model (MLRM) agents to answer complex visual queries by repeatedly inspecting image regions, gro

Visual Graph Scaffolds for Structural Reasoning in Large Language Models

TutorialsDGX agent

arXiv:2606.02673v1 Announce Type: new Abstract: Graphs have been used to enhance large language models (LLMs) for structured reasoning, mostly as external knowledge sources are provided to models at t

Content type
AllBlogX PostPaperYouTubeRedditGitHub

Visual Instruction Tuning Aligns Modalities through Abstraction

Local AiDGX agent

arXiv:2606.03871v1 Announce Type: cross Abstract: Visual instruction tuning effectively adapts a pre-trained Large Language Model (LLM) to process image information alongside text. Yet, it remains unc

VLA-Arena: An Open-Source Framework for Benchmarking Vision-Language-Action Models

Model ReleasesDGX agent

arXiv:2512.22539v2 Announce Type: replace-cross Abstract: While Vision-Language-Action models (VLAs) are rapidly advancing towards generalist robot policies, it remains difficult to quantitatively und

VLESA: Vision-Language Embodied Safety Agent for Human Activity Monitoring

Model ReleasesDGX agent

arXiv:2606.03954v1 Announce Type: new Abstract: As AI systems increasingly assist humans in physical tasks, ensuring safety becomes paramount -- physical actions carry immediate and irreversible conse

vLLM Semantic Router: Signal Driven Decision Routing for Mixture-of-Modality Models

Model ReleasesDGX agent

arXiv:2603.04444v3 Announce Type: replace-cross Abstract: As large language models (LLMs) diversify across modalities, capabilities, and cost profiles, the problem of intelligent request routing -- se

VulnAgent-R2: Evidence-Calibrated Multi-Agent Auditing for Repository-Level Vulnerability Detection

Local AiDGX agent

arXiv:2603.13384v2 Announce Type: replace-cross Abstract: Software vulnerabilities often depend on cross-file data flow, build options, framework conventions, and runtime guards, so isolated function

WaterSIC: Information-Theoretically (Near) Optimal Linear Layer Quantization

Model ReleasesDGX agent

arXiv:2603.04956v2 Announce Type: replace Abstract: This paper considers the problem of converting a given dense linear layer to low precision. The tradeoff between compressed length and output discre

Wavelet as Tokenizer: Preliminary Results on a Shared Wavelet Token Schema for Natural Signals

ResearchDGX agent

arXiv:2606.02631v1 Announce Type: cross Abstract: This paper studies whether audio, images, and video can share a common wavelet token schema rather than relying on separate modality-specific latent g

Wavelet Fourier Diffuser: Frequency-Aware Diffusion Model for Reinforcement Learning

Model ReleasesDGX agent

arXiv:2509.19305v2 Announce Type: replace-cross Abstract: Diffusion probability models have shown significant promise in offline reinforcement learning by directly modeling trajectory sequences. Howev

we are seeing costs start to matter! uber just set limits of $1500 in tokens per developer per month i think we're going to start seeing mor…

AgentsDGX agent

we are seeing costs start to matter! uber just set limits of $1500 in tokens per developer per month i think we're going to start seeing more of this, and LangSmith Gateway is a great way to implement

We have a winner of the Lenny's Newsletter x @Replit buildathon 🏆 Congratulations @andrey_esipov on your mind-blowing project, Operators. Y…

ToolsDGX agent

We have a winner of the Lenny's Newsletter x @Replit buildathon 🏆 Congratulations @andrey_esipov on your mind-blowing project, Operators. Y'all need to check this out: https://operators--andreyyesipov

We have partnered with @Cognition to bring their AI coding platform, Devin, to the Public Sector. Together, we’re helping organizations mode…

TutorialsDGX agent

We have partnered with @Cognition to bring their AI coding platform, Devin, to the Public Sector. Together, we’re helping organizations modernize software development and accelerate secure AI adoption

We just announced the integration of our knowledge engine, Pinecone Nexus, with @Microsoft OneLake at #MSBuild. Want to build a reliable, pr…

ApplicationsDGX agent

We just announced the integration of our knowledge engine, Pinecone Nexus, with @Microsoft OneLake at #MSBuild. Want to build a reliable, production-grade knowledge layer directly over your structured

We partnered with @FireworksAI_HQ to train open-source models for legal. Here's what we found: 1) Hybrid legal agents can beat frontier mode…

Model ReleasesDGX agent

We partnered with @FireworksAI_HQ to train open-source models for legal. Here's what we found: 1) Hybrid legal agents can beat frontier models on quality and cost by routing selectively to a frontier

Weak Diffusion Priors Can Still Achieve Strong Inverse-Problem Performance

ResearchDGX agent

arXiv:2601.22443v2 Announce Type: replace-cross Abstract: Can a diffusion model trained on bedrooms recover human faces? Diffusion models are widely used as priors for inverse problems, but standard a

WebRISE: Requirement-Induced State Evaluation for MLLM-Generated Web Artifacts

Local AiDGX agent

arXiv:2606.03220v1 Announce Type: cross Abstract: Existing benchmarks for MLLM-generated web artifacts assess interaction through local evidence and miss the requirement-induced states and transitions

Welp, that happened faster than I predicted. Thought it would be end of 2027, then early 2027, but agentic traffic growing so fast that bots…

AgentsDGX agent

Welp, that happened faster than I predicted. Thought it would be end of 2027, then early 2027, but agentic traffic growing so fast that bots have now passed human traffic online for the first time in

We’re bringing new capabilities to GPT-Rosalind, a model series purpose-built for life sciences research at enterprise scale. It brings GPT-…

Model ReleasesDGX agent

We’re bringing new capabilities to GPT-Rosalind, a model series purpose-built for life sciences research at enterprise scale. It brings GPT-5.5’s agentic coding and tool use together with stronger int

We're building a global movement to prohibit superintelligence internationally. Our new campaign in Canada is supported by over 30 MPs and S…

SafetyDGX agent

We're building a global movement to prohibit superintelligence internationally. Our new campaign in Canada is supported by over 30 MPs and Senators calling for a ban. After our success in the UK, it's

What are you investing in? Hopium:

SafetyDGX agent

What are you investing in? Hopium: Former BlackRock fund manager Ed Dowd on the stock market: 'If 45% of your market cap is AI and there's no profits yet, what are you investing in?' 'You're investing

What Benchmarks Don't Measure: The Case for Evaluating Abstention Competence in Autonomous Agents

Model ReleasesDGX agent

arXiv:2606.02965v1 Announce Type: new Abstract: Benchmarks for autonomous agents measure whether agents complete tasks, yet this framing is systematically blind to whether an agent should have proceed

What Do Students Learn? A Feature-Level Analysis of Dark Knowledge

TutorialsDGX agent

arXiv:2606.03052v1 Announce Type: new Abstract: Knowledge Distillation (KD) is a powerful tool for model compression, yet the precise mechanisms by which student models acquire feature representations

What Makes Interaction Trajectories Effective for Training Terminal Agents?

Model ReleasesDGX agent

arXiv:2606.03461v1 Announce Type: new Abstract: Stronger code agents are commonly assumed to be superior teachers for post-training, yet this assumption remains poorly disentangled from task difficult

What’s new in serverless Managed Service for Apache Spark

HardwareDGX agent

Whether you use it for data preparation, real-time interactive queries, AI model training, or something entirely different, running Apache Spark at scale is demanding — you shouldn’t have to manage th

What's the most unhinged thing you've used an uncensored Ollama model for? Also... what are the best uncensored models right now?

Local AiDGX agent

This Reddit discussion from r/ollama explores user experiences with uncensored Ollama language models, featuring anecdotal accounts of unusual or extreme use cases and recommendations for popular unce

Wheel-Mounted/GNSS Fusion with AI-Aided Position Updates

AgentsDGX agent

arXiv:2606.03265v1 Announce Type: new Abstract: Accurate and robust localization remains a fundamental challenge for autonomous ground vehicles. In this work, we propose a hybrid neural inertial navig

When Attention Collapses: Stage-Aware Visual Token Pruning from Structure to Semantics

SafetyDGX agent

arXiv:2606.03569v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) have demonstrated remarkable capabilities but suffer from significant computational overhead during inference. While vis

When Does Complexity Conditioning Help a Frozen Sentence Embedding? A Controlled Study of Per-Sentence and Pair-Level Difficulty Adaptation

ResearchDGX agent

arXiv:2606.03244v1 Announce Type: new Abstract: A common intuition is that sentence embeddings should adapt to the difficulty of the input. We test this intuition in a controlled, multi-seed setting:

When Graph Tokens Sink: A Mechanistic Analysis of Graph Language Models

SafetyDGX agent

arXiv:2606.03712v1 Announce Type: new Abstract: Graph Language Models (GLMs) have become a promising direction for adapting Large Language Models (LLMs) to graph learning tasks. By transforming graph

When Helping Hurts and How to Fix It: Multi-Agent Debate for Data Cleaning

AgentsDGX agent

arXiv:2606.02866v1 Announce Type: new Abstract: When does multi-agent debate help data cleaning, and when does it hurt? Across three benchmarks, four model families, and over 6,000 task-condition pair

When Model Merging Breaks Routing: Training-Free Calibration for MoE

Model ReleasesDGX agent

arXiv:2606.03391v1 Announce Type: cross Abstract: Model merging has emerged as a cost-effective approach for consolidating the capabilities of multiple LLMs without retraining. However, existing mergi

When Models Refuse: Political Steerability and Feature Richness as Measures of Ideological Depth

SafetyDGX agent

arXiv:2508.21448v3 Announce Type: replace Abstract: Large language models (LLMs) sometimes refuse to follow benign instructions, such as declining to argue a political position or adopt a stated perso

When RLHF Fails: A Mechanistic Taxonomy of Reward Hacking, Collapse, and Evaluator Gaming

Local AiDGX agent

arXiv:2606.03238v1 Announce Type: cross Abstract: Reinforcement learning from human feedback (RLHF) makes large-scale post-training possible by replacing an underspecified human objective with learned

when sam quotes the bible, you know things aren’t going well for OpenAI

SafetyDGX agent

when sam quotes the bible, you know things aren’t going well for OpenAI one of the quotes i find most inspiring on a hard day: 'Whatever your hand finds to do, do it with all your might, for in the re

When Should LLMs Be Less Specific? Selective Abstraction for Reliable Long-Form Text Generation

ResearchDGX agent

arXiv:2602.11908v3 Announce Type: replace Abstract: LLMs are widely used, yet they remain prone to factual errors that erode user trust and limit adoption in high-risk settings. One approach to mitiga

When Should the Teacher Move? Temporal Coupling and Stability in Self On-Policy Distillation

Model ReleasesDGX agent

arXiv:2606.03532v1 Announce Type: cross Abstract: Self on-policy distillation trains a student policy against a teacher derived from its own parameter history, yet the teacher's update schedule -- whi

When the country needed a serious principled political party on the “right”, I launched @_AdvanceUK. The situation has since changed. And so…

IndustryDGX agent

When the country needed a serious principled political party on the “right”, I launched @_AdvanceUK. The situation has since changed. And so we too must change. We are not in politics for the sake of

When to Re-Plan: Subgoal Persistence in Hierarchical Latent Reasoning

SafetyDGX agent

arXiv:2606.03741v1 Announce Type: new Abstract: Long-horizon reasoning requires a system to commit to medium-horizon intent without becoming rigid: re-plan too often and computation never coheres into

When you paste a prompt from your clipboard, but it turns out to be a text message

IndustryDGX agent

This post likely humorously describes the awkward situation where a user intends to paste a saved AI prompt into a tool or application, but accidentally pastes a text message from their clipboard inst

Where Do We (Not) Need Temporal Context in Low-Resource Video Task Adaptation?

Model ReleasesDGX agent

arXiv:2606.03837v1 Announce Type: new Abstract: Parameter-efficient fine-tuning (PEFT) and probing enable adaptation of foundation models using only a small number of trainable parameters, making it a

Which Defense Closes Which Threat? Attributing OWASP-LLM-Top-10 Coverage and Its Brittleness Under Paraphrasing

Model ReleasesDGX agent

arXiv:2606.02822v1 Announce Type: cross Abstract: Production LLM applications stack several defense families -- refusal-phrase filters, token-budget controls, model allowlists, rate limits, tool-regis

White Girls Raped By Dogs, Whisky Bottles, & 100s Of Men: Britain's Migrant Grooming-Gang Scandal Exposed https://www.zerohedge.com/politica…

IndustryDGX agent

White Girls Raped By Dogs, Whisky Bottles, & 100s Of Men: Britain's Migrant Grooming-Gang Scandal Exposed https://www.zerohedge.com/political/white-girls-raped-dogs-whisky-bottles-100s-men-britains-mi

Who Deserves the Reward? SHARP: Shapley Credit-based Optimization for Multi-Agent System

SafetyDGX agent

arXiv:2602.08335v2 Announce Type: replace Abstract: Integrating Large Language Models (LLMs) with external tools via multi-agent systems offers a promising new paradigm for decomposing and solving com

Who has used ChatGPT or any other AI for coding without knowing how to code?

TutorialsDGX agent

A Reddit discussion from r/ChatGPT where users share experiences using AI tools like ChatGPT to write code despite lacking programming knowledge or expertise. The thread likely documents how individua

Whom to Query for What: Adaptive Group Elicitation via Multi-Turn LLM Interactions

AgentsDGX agent

arXiv:2602.14279v2 Announce Type: replace-cross Abstract: Eliciting information to reduce uncertainty about latent group-level properties from surveys and other collective assessments requires allocat

Whoop is building agentic AI maturity on a foundation of enterprise health data

AgentsDGX agent

Enterprise AI programs are moving beyond experimentation, but agentic AI maturity — the ability to run governed, autonomous workflows at production scale — remains out of reach for most organizations.

Whose Name Comes Up? II: Benchmarking and Intervention-Based Auditing of LLM-Based Scholar Recommendation

Model ReleasesDGX agent

arXiv:2602.08873v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are now used for academic expert recommendation. Existing audits typically evaluate such recommendations in isola

Why Are Linear RNNs More Parallelizable?

ResearchDGX agent

arXiv:2603.03612v3 Announce Type: replace-cross Abstract: The community is increasingly exploring linear RNNs (LRNNs) as language models, motivated by their expressive power and parallelizability. Whi

Will Accurate Fields Mislead Photonic Design? FromGlobal Accuracy to Port Readout

Model ReleasesDGX agent

arXiv:2606.03038v1 Announce Type: new Abstract: Neural field surrogates can accelerate photonic design loops, but a surrogate that looks accurate in global field error can still mis-rank candidate dev

WISE: A World Knowledge-Informed Semantic Evaluation for Text-to-Image Generation

Model ReleasesDGX agent

arXiv:2503.07265v4 Announce Type: replace-cross Abstract: Text-to-Image (T2I) models are capable of generating high-quality artistic creations and visual content. However, existing research and evalua

WISE-HAR: A Generalizable Ensemble Deep Learning Framework for WiFi-Based Human Activity Recognition

ApplicationsDGX agent

arXiv:2606.02974v1 Announce Type: new Abstract: Human Activity Recognition (HAR) using WiFi signals has emerged as a transformative technology for smart homes, healthcare monitoring, security systems,

With LangSmith Engine, systemic issues get surfaced automatically instead of getting buried in traces. @ollieelmgren from @ListenLabs on how…

AgentsDGX agent

With LangSmith Engine, systemic issues get surfaced automatically instead of getting buried in traces. @ollieelmgren from @ListenLabs on how LangSmith Engine changed the way his team evaluates their a

With new Majorana 2 quantum chip, Microsoft claims dramatic breakthrough in qubit stability

IndustryDGX agent

Microsoft Corp. today announced an updated chip for quantum computing called Majorana 2 that it says is 1,000 times more reliable in terms of qubit stability. According to the company, it opens the do

Wordle 1,809 5/6 🟨⬛🟨⬛⬛ ⬛🟨⬛⬛⬛ ⬛🟩🟨🟨⬛ 🟩🟩🟩🟩⬛ 🟩🟩🟩🟩🟩

Model ReleasesDGX agent

This post documents a completed Wordle puzzle (puzzle #1,809) solved in five attempts, showing the progression of letter guesses through color-coded feedback (yellow for correct letters in wrong posit

Working with @FireworksAI_HQ to make MAI models easy to fine-tune and fully yours.

TutorialsDGX agent

Working with @FireworksAI_HQ to make MAI models easy to fine-tune and fully yours. Microsoft MAI models. Coming soon to Fireworks. Intelligence you control. End-to-end lineage you can prove. Fine-tune

World Models Meet Language Models: On the Complementarity of Concrete and Abstract Reasoning

SafetyDGX agent

arXiv:2606.03603v1 Announce Type: cross Abstract: World models and multimodal large language models (MLLMs) provide complementary capabilities for predicting future outcomes from static visual observa

Worth Remembering: Surprise-Gated Robot Episodic Memory

ResearchDGX agent

arXiv:2606.03787v1 Announce Type: new Abstract: Robots solving generalist tasks need to be able to ground instructions in their past experience, since humans may refer to notable past events when givi

Wow, the free game sucked my Netlify credits dry in an hour. Re-deploying a more bandwidth friendly version (your saved game should still wo…

ApplicationsDGX agent

Ethan Mollick shared that a free game he deployed on Netlify consumed his monthly bandwidth credits in just one hour, forcing him to redeploy an optimized, more bandwidth-efficient version while prese

← Previous
1…726727728729730…1518
Next →