AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,606Total entries
1Added by human
84,605Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
84,606 results
Model Releases

VLM-3R: Vision-Language Models Augmented with Instruction-Aligned 3D Reconstruction

DGX agent

arXiv:2505.20279v4 Announce Type: replace-cross Abstract: The rapid advancement of Large Multimodal Models (LMMs) for 2D images and videos has motivated extending these models to understand 3D scenes,

model-releasesarxiv-cs-cl
21 Apr 2026
Research
X Post
Paper
YouTube
Reddit
GitHub

Vocab Diet: Reshaping the Vocabulary of LLMs via Vector Arithmetic

DGX agent

arXiv:2510.17001v2 Announce Type: replace Abstract: Large language models (LLMs) often encode word-form variation (e.g., walk vs. walked) as linear directions in the embedding space. However, standard

researcharxiv-cs-cl
21 Apr 2026
Local Ai

VocabTailor: Dynamic Vocabulary Selection for Downstream Tasks in Small Language Models

DGX agent

arXiv:2508.15229v3 Announce Type: replace Abstract: Small Language Models (SLMs) provide computational advantages in resource-constrained environments, yet memory limitations remain a critical bottlen

local-aiarxiv-cs-cl
21 Apr 2026
Model Releases

Voronoi-guided Bilateral 2D Gaussian Splatting for Arbitrary-Scale Hyperspectral Image Super-Resolution

DGX agent

arXiv:2604.17727v1 Announce Type: new Abstract: Most existing hyperspectral image super-resolution methods require modifications for different scales, limiting their flexibility in arbitrary-scale rec

model-releasesarxiv-cs-cv
21 Apr 2026
Safety

Waking Up Blind: Cold-Start Optimization of Supervision-Free Agentic Trajectories for Grounded Visual Perception

DGX agent

arXiv:2604.17475v1 Announce Type: cross Abstract: Small Vision-Language Models (SVLMs) are efficient task controllers but often suffer from visual brittleness and poor tool orchestration. They typical

safetyarxiv-cs-cl
21 Apr 2026
Research

Wasserstein Distributionally Robust Risk-Sensitive Estimation via Conditional Value-at-Risk

DGX agent

arXiv:2604.18546v1 Announce Type: new Abstract: We propose a distributionally robust approach to risk-sensitive estimation of an unknown signal x from an observed signal y. The unknown signal and obse

researcharxiv-cs-lg
21 Apr 2026
Research

Wasserstein-p Central Limit Theorem Rates: From Local Dependence to Markov Chains

DGX agent

arXiv:2601.08184v3 Announce Type: replace-cross Abstract: Non-asymptotic central limit theorem (CLT) rates play a central role in modern machine learning and operations research. In this paper, we stu

researcharxiv-cs-lg
21 Apr 2026
Agents

We are entering an extremely exciting era for open-weight models. Kimi K2.6 now feels like a top agentic model. I took it for a spin via @Fi…

DGX agent

We are entering an extremely exciting era for open-weight models. Kimi K2.6 now feels like a top agentic model. I took it for a spin via @FireworksAI_HQ fast inference APIs. Kimi K2.6 has impressive a

agentsdair-ai--x
21 Apr 2026
Applications

We are excited to have @FireworksAI_HQ as a day 0 launch partner for Kimi K2.6! Their inference and fine-tuning platform is fast, reliable, …

DGX agent

We are excited to have @FireworksAI_HQ as a day 0 launch partner for Kimi K2.6! Their inference and fine-tuning platform is fast, reliable, and scales well under real production load, making it easy t

applicationskimi-moonshot--x
21 Apr 2026
Safety

We think ControlAI can turn $50M / year into a 10% chance of banning ASI. Most of the AI safety community has been far too coy about extinct…

DGX agent

We think ControlAI can turn $50M / year into a 10% chance of banning ASI. Most of the AI safety community has been far too coy about extinction risk. We're not. It's not that complicated: AI smarter t

safetyconnor-leahy--x
21 Apr 2026
Safety

Weakly-Supervised Referring Video Object Segmentation through Text Supervision

DGX agent

arXiv:2604.17797v1 Announce Type: new Abstract: Referring video object segmentation (RVOS) aims to segment the target instance in a video, referred by a text expression. Conventional approaches are mo

safetyarxiv-cs-cv
21 Apr 2026
Model Releases

WeatherArchive-Bench: Benchmarking Retrieval-Augmented Reasoning for Historical Weather Archives

DGX agent

arXiv:2510.05336v2 Announce Type: replace Abstract: Historical archives on weather events are collections of enduring primary source records that offer rich, untapped narratives of how societies have

model-releasesarxiv-cs-cl
21 Apr 2026
Hardware

Web-Gewu: A Browser-Based Interactive Playground for Robot Reinforcement Learning

DGX agent

arXiv:2604.17050v1 Announce Type: new Abstract: With the rapid development of embodied intelligence, robotics education faces a dual challenge: high computational barriers and cumbersome environment c

hardwarearxiv-cs-ro
21 Apr 2026
Model Releases

🔴 We're going live tomorrow for another Come Build with Pinecone session! Join us to build, break, and learn together. Each session, we pic…

DGX agent

🔴 We're going live tomorrow for another Come Build with Pinecone session! Join us to build, break, and learn together. Each session, we pick up a real project — building RAG pipelines, wiring up agent

model-releasespinecone--x
21 Apr 2026
Applications

We're honored to be named Google Cloud's 2026 AI Tooling Partner of the Year. This recognition reflects what 50 million builders have made p…

DGX agent

We're honored to be named Google Cloud's 2026 AI Tooling Partner of the Year. This recognition reflects what 50 million builders have made possible together. Product managers, founders, students, oper

applicationsreplit--x
21 Apr 2026
Model Releases

We're open-sourcing FlashKDA — our high-performance CUTLASS-based implementation of Kimi Delta Attention kernels. Achieves 1.72×–2.22× prefi…

DGX agent

We're open-sourcing FlashKDA — our high-performance CUTLASS-based implementation of Kimi Delta Attention kernels. Achieves 1.72×–2.22× prefill speedup over the flash-linear-attention baseline on H20,

model-releaseskimi-moonshot--x
21 Apr 2026
Tools

We're partnering with SpaceX to improve Composer. http://cursor.com/blog/spacex-model-training

DGX agent

Cursor announced a partnership with SpaceX to enhance their Composer feature, likely involving improved AI model training capabilities or computational resources. The collaboration suggests integratin

toolscursor--x
21 Apr 2026
Tools

We've reduced memory crashes in the Cursor desktop application by 80% since February. Here's how we detect, debug, and prevent OOMs at scale…

DGX agent

Cursor reduced memory crashes in its desktop application by 80% since February through improvements in out-of-memory (OOM) detection, debugging, and prevention at scale. The post likely details the te

toolscursor--x
21 Apr 2026
Local Ai

What are you guys using to train LTX 2.3 loras locally on 4090s?

DGX agent

Users training LTX-2.3 LoRAs on RTX 4090s typically use the official ltx-trainer tool, though the model officially targets H100 GPUs with lower VRAM setups requiring gradient checkpointing and reduced

local-air-stablediffusion
21 Apr 2026
Research

What If Consensus Lies? Selective-Complementary Reinforcement Learning at Test Time

DGX agent

arXiv:2603.19880v2 Announce Type: replace Abstract: Test-Time Reinforcement Learning (TTRL) enables Large Language Models (LLMs) to enhance reasoning capabilities on unlabeled test streams by deriving

researcharxiv-cs-lg
21 Apr 2026
Research

What is born of light is light

DGX agent

This post likely discusses how light or illumination (literal or metaphorical) generates or produces similar qualities, potentially referencing philosophical, scientific, or spiritual concepts about t

researchnous-research--x
21 Apr 2026
Agents

What Makes AI Research Replicable? Executable Knowledge Graphs as Scientific Knowledge Representations

DGX agent

arXiv:2510.17795v3 Announce Type: replace Abstract: Replicating AI research is a crucial yet challenging task for large language model (LLM) agents. Existing approaches often struggle to generate exec

agentsarxiv-cs-cl
21 Apr 2026
Research

What makes an entity salient in discourse?

DGX agent

arXiv:2508.16464v2 Announce Type: replace Abstract: Entities in discourse vary in salience: main participants, objects and locations stay prominent, while others are quickly forgotten, raising questio

researcharxiv-cs-cl
21 Apr 2026
Model Releases

What makes ChatGPT Images 2.0 a state-of-the-art image generation model? Researchers behind the model explain. A thread: Thinking & Intellig…

DGX agent

What makes ChatGPT Images 2.0 a state-of-the-art image generation model? Researchers behind the model explain. A thread: Thinking & Intelligence in ChatGPT Images 2.0, demonstrated by @ayaanzhaque Med

model-releasesopenai--x
21 Apr 2026
Industry

What to expect during Appian World: Join theCUBE April 27-29

DGX agent

Real value is emerging as process-centric AI becomes embedded in how work actually gets done. Enterprises are moving beyond isolated automation and embedding AI directly into workflows where decisions

industrysiliconangle
21 Apr 2026
Local Ai

What’s everyone’s favorite sampler and scheduler these days?

DGX agent

This Reddit discussion explores sampler and scheduler preferences within the Stable Diffusion community, where samplers guide the process of turning noise into an image and schedulers control how nois

local-air-stablediffusion
21 Apr 2026
Model Releases

What's Left Unsaid? Detecting and Correcting Misleading Omissions in Multimodal News Previews

DGX agent

arXiv:2601.05563v2 Announce Type: replace Abstract: Even when factually correct, social-media news previews (image-headline pairs) can induce interpretation drift: by selectively omitting crucial cont

model-releasesarxiv-cs-cv
21 Apr 2026
Tools

What’s next in tech? Come find out 👀 @Replit is heading to Berkeley for a demo of our newest features and talent search for internships and…

DGX agent

What’s next in tech? Come find out 👀 @Replit is heading to Berkeley for a demo of our newest features and talent search for internships and new grad roles. We will also be handing out our exclusive me

toolsreplit--x
21 Apr 2026
Industry

What's the deal with spacesuits for the Moon? Will they be ready in time?

DGX agent

NASA's next-generation spacesuit for the Artemis III mission, the AxEMU developed by Axiom Space, has passed a contractor-led technical review and continues undergoing testing with NASA astronauts and

industryars-technica
21 Apr 2026
Research

When Background Matters: Breaking Medical Vision Language Models by Transferable Attack

DGX agent

arXiv:2604.17318v1 Announce Type: new Abstract: Vision-Language Models (VLMs) are increasingly used in clinical diagnostics, yet their robustness to adversarial attacks remains largely unexplored, pos

researcharxiv-cs-cv
21 Apr 2026
Tutorials

When Can LLMs Learn to Reason with Weak Supervision?

DGX agent

arXiv:2604.18574v1 Announce Type: new Abstract: Large language models have achieved significant reasoning improvements through reinforcement learning with verifiable rewards (RLVR). Yet as model capab

tutorialsarxiv-cs-lg
21 Apr 2026
Safety

When Choices Become Risks: Safety Failures of Large Language Models under Multiple-Choice Constraints

DGX agent

arXiv:2604.16916v1 Announce Type: new Abstract: Safety alignment in large language models (LLMs) is primarily evaluated under open-ended generation, where models can mitigate risk by refusing to respo

safetyarxiv-cs-cl
21 Apr 2026
Model Releases

When Earth Foundation Models Meet Diffusion: An Application to Land Surface Temperature Super-Resolution

DGX agent

arXiv:2604.16841v1 Announce Type: new Abstract: Land surface temperature (LST) super-resolution is important for environmental monitoring. However, it remains challenging as coarse thermal observation

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

When Helpers Become Hazards: A Benchmark for Analyzing Multimodal LLM-Powered Safety in Daily Life

DGX agent

arXiv:2601.04043v2 Announce Type: replace Abstract: As Multimodal Large Language Models (MLLMs) become an indispensable assistant in human life, the unsafe content generated by MLLMs poses a danger to

model-releasesarxiv-cs-cl
21 Apr 2026
Research

When Informal Text Breaks NLI: Tokenization Failure, Distribution Shift, and Targeted Mitigations

DGX agent

arXiv:2604.16787v1 Announce Type: new Abstract: We study how informal surface forms degrade NLI accuracy in ELECTRA-small (14M) and RoBERTa-large (355M) across four transforms applied to SNLI and Mult

researcharxiv-cs-cl
21 Apr 2026
Research

When Misinformation Speaks and Converses: Rethinking Fact-Checking in Audio Platforms

DGX agent

arXiv:2604.16767v1 Announce Type: new Abstract: Audio platforms have evolved beyond entertainment. They have become central to public discourse, from podcasts and radio to WhatsApp voice notes and liv

researcharxiv-cs-cl
21 Apr 2026
Research

When More Words Say Less: Decoupling Length and Specificity in Image Description Evaluation

DGX agent

arXiv:2601.04609v2 Announce Type: replace Abstract: Vision-language models (VLMs) are increasingly used to make visual content accessible via text-based descriptions. In current systems, however, desc

researcharxiv-cs-cl
21 Apr 2026
Model Releases

When Pretty Isn't Useful: Investigating Why Modern Text-to-Image Models Fail as Reliable Training Data Generators

DGX agent

arXiv:2602.19946v4 Announce Type: replace Abstract: Recent text-to-image (T2I) diffusion models produce visually stunning images and demonstrate excellent prompt following. But do they perform well as

model-releasesarxiv-cs-cv
21 Apr 2026
Research

When Seeing Overrides Knowing: Disentangling Knowledge Conflicts in Vision-Language Models

DGX agent

arXiv:2507.13868v2 Announce Type: replace Abstract: Vision-language models (VLMs) increasingly combine visual and textual information to perform complex tasks. However, conflicts between their interna

researcharxiv-cs-cv
21 Apr 2026
Model Releases

When Spike Sparsity Does Not Translate to Deployed Cost: VS-WNO on Jetson Orin Nano

DGX agent

arXiv:2604.17040v1 Announce Type: new Abstract: Spiking neural operators are appealing for neuromorphic edge computing because event-driven substrates can, in principle, translate sparse activity into

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

When Text Hijacks Vision: Benchmarking and Mitigating Text Overlay-Induced Hallucination in Vision Language Models

DGX agent

arXiv:2604.17375v1 Announce Type: new Abstract: Recent advances in Vision-Language Models (VLMs) have substantially enhanced their ability across multimodal video understanding benchmarks spanning tem

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

When Visuals Aren't the Problem: Evaluating Vision-Language Models on Misleading Data Visualizations

DGX agent

arXiv:2603.22368v2 Announce Type: replace Abstract: Visualizations help communicate data insights, but deceptive data representations can distort their interpretation and propagate misinformation. Whi

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

When W4A4 Breaks Camouflaged Object Detection: Token-Group Dual-Constraint Activation Quantization

DGX agent

arXiv:2604.16855v1 Announce Type: new Abstract: Camouflaged object detection (COD) segments objects that intentionally blend with the background, so predictions depend on subtle texture and boundary c

model-releasesarxiv-cs-cv
21 Apr 2026
Safety

Where Do Self-Supervised Speech Models Become Unfair?

DGX agent

arXiv:2604.18249v1 Announce Type: new Abstract: Speech encoder models are known to model members of some speaker groups (SGs) better than others. However, there has been little work in establishing wh

safetyarxiv-cs-cl
21 Apr 2026
Research

Where is the Mind? Persona Vectors and LLM Individuation

DGX agent

arXiv:2604.17031v1 Announce Type: new Abstract: The individuation problem for large language models asks which entities associated with them, if any, should be identified as minds. We approach this pr

researcharxiv-cs-cl
21 Apr 2026
Local Ai

Where to Focus: Query-Modulated Multimodal Keyframe Selection for Long Video Understanding

DGX agent

arXiv:2604.17422v1 Announce Type: new Abstract: Long video understanding remains a formidable challenge for Multimodal Large Language Models (MLLMs) due to the prohibitive computational cost of proces

local-aiarxiv-cs-cv
21 Apr 2026
Model Releases

Where's the raccoon with the ham radio? (ChatGPT Images 2.0)

DGX agent

OpenAI released ChatGPT Images 2.0 today, their latest image generation model. On the livestream Sam Altman said that the leap from gpt-image-1 to gpt-image-2 was equivalent to jumping from GPT-3 to G

model-releasessimon-willison
21 Apr 2026
Agents

Whistant: A Standalone AI Agent for iPhone — No Mac Required

DGX agent

Whistant is an on-phone AI buddy that helps users get things done directly from their iPhone without requiring a Mac. It breaks requests into step-by-step subtasks and executes them on the device, inc

agentsr-ollama
21 Apr 2026
← Previous
1…15601561156215631564…1763
Next →