AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
Human
91,060Total entries
1Added by human
91,059Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
91,059 results
5 Jun 2026

GLASS: GRPO-Trained LoRA for Acoustic Style Steering in Zero-Shot Text-to-Speech

SafetyDGX agent

arXiv:2606.05889v1 Announce Type: cross Abstract: We propose GLASS, a framework for composable acoustic style control in zero-shot autoregressive text-to-speech (TTS) that learns controls from post-ge

Global Cross-Modal Geo-Localization: A Million-Scale Dataset and a Physical Consistency Learning Framework

Model ReleasesDGX agent

arXiv:2603.08491v2 Announce Type: replace Abstract: Cross-modal Geo-localization (CMGL) matches ground-level text descriptions with geo-tagged aerial imagery, which is crucial for pedestrian navigatio

Global-Local Monte Carlo Tree Search in Vision-Language Models for Text-to-3D Indoor Scene Generation

ResearchDGX agent
DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2606.06002v1 Announce Type: new Abstract: Large Vision-Language Models have achieved significant reasoning performance in various tasks.However, there are few studies on text-to-3D indoor scene

GMBFormer: An NDVI-Guided Global Memory Bank Transformer for Urban Green-Space Extraction from Ultra-High-Resolution Imagery

ResearchDGX agent

arXiv:2606.06363v1 Announce Type: new Abstract: Urban green-space extraction from ultra-high-resolution (UHR) imagery is commonly performed patch by patch, which limits semantic reuse among spatially

Google Cloud says its SpaceX compute deal is a 'short-term' agreement 'to ensure we have bridge capacity to meet surging customer demand' for Gemini Enterprise (Kate Conger/New York Times)

Model ReleasesDGX agent

Kate Conger / New York Times: Google Cloud says its SpaceX compute deal is a “short-term” agreement “to ensure we have bridge capacity to meet surging customer demand” for Gemini Enterprise — Elon Mus

Gotta Grow Fast: Design and Benchmarking of a Tip Mount for High-Speed Vine Robots

HardwareDGX agent

arXiv:2606.06040v1 Announce Type: new Abstract: Soft, growing vine robots extend through tip eversion, a mechanism that enables navigation through cluttered environments. However, integrating cameras

GRAMformer: Any-Order Modality Interactions via Volumetric Multimodal Cross-Attention

ResearchDGX agent

arXiv:2606.06249v1 Announce Type: new Abstract: Transformer-based multimodal models rely on attention mechanisms to integrate information across heterogeneous modalities. Despite their success, existi

Grok Build now operates within your project files. Full read and write plus all new files end up in the project. Big improvement for enterpr…

ApplicationsDGX agent

Grok Build now operates within your project files. Full read and write plus all new files end up in the project. Big improvement for enterprise workflow creation. You can simply type 'work within my p

Grok Build updates

IndustryDGX agent

Grok Build updates Bug fixes shipping to Grok Build 0.2.20 (release notes will be available in the TUI and on change-log website) • Eliminate ghost-cell artifacts in markdown table rendering • Make mo

Grok model improvement

AgentsDGX agent

Grok model improvement The updated Grok-build model (still the 0.5T one) is much better than before. It’s less lazy, more autonomous, and more accurate. We are still improving it on long-horizon tasks

Grok supports worktrees

IndustryDGX agent

Grok supports worktrees Grok Build tip of the day: worktrees! If you're unfamiliar with worktrees, they're essentially lightweight copies of your repo, allowing you to run parallel agents within their

Grounded but Misleading: Evaluating Semantic Alignment in AI-Generated Security Explanations

SafetyDGX agent

arXiv:2602.05056v2 Announce Type: replace-cross Abstract: Online scams increasingly leverage fluent and context-aware social engineering strategies, creating growing demand for AI systems that explain

GS-NFS: Bandwidth-adaptive Streaming of Dynamic Gaussian Splats and Point Clouds

HardwareDGX agent

arXiv:2606.05650v1 Announce Type: cross Abstract: Dynamic 3D Gaussian Splatting (3DGS) holds great promise as a 3D video streaming technology since it can represent complex 3D scenes with high fidelit

HANDOFF: Humanoid Agentic Task-Space Whole-Body Control via Distilled Complementary Teachers

SafetyDGX agent

arXiv:2606.06493v1 Announce Type: new Abstract: For a humanoid robot to be deployed in the real world, the choice of command space (i.e., the interface between task planning and whole-body control) is

Harmonious Parameter Adaptation in Continual Visual Instruction Tuning for Safety-Aligned MLLMs

Model ReleasesDGX agent

arXiv:2511.20158v2 Announce Type: replace Abstract: While continual visual instruction tuning (CVIT) has shown promise in adapting multimodal large language models (MLLMs), existing studies predominan

Harnessing Generalist Agents for Contextualized Time Series

AgentsDGX agent

arXiv:2606.05404v1 Announce Type: cross Abstract: Time series are often embedded in rich contexts that are essential for holistic modeling. Moreover, real-world practitioners often require end-to-end

Harnessing Structural Context for Entity Alignment Foundation Models

Model ReleasesDGX agent

arXiv:2606.06109v1 Announce Type: new Abstract: Entity alignment (EA) aims to identify equivalent entities across heterogeneous knowledge graphs (KGs) and is a key component of knowledge fusion and cr

Have you tried the new Replit Canvas? - Create beautiful UI designs with AI - Generate assets with GPT-Image 2 & Seedance - Turn your design…

ToolsDGX agent

Replit Canvas is a new AI-powered design tool that enables users to create beautiful UI designs with artificial intelligence assistance, including GPT-Image 2 for image generation and Seedance for ass

HDST-GNN: Heterogeneous Dynamic Spatiotemporal Graph Neural Networks for Multi-Object Tracking in UAV Aerial Imagery

ResearchDGX agent

arXiv:2606.05587v1 Announce Type: new Abstract: Multi-object tracking (MOT) from UAV imagery presents unique challenges: altitude varies across sequences, objects are small and densely packed, and fre

Henry Nowak died the same way a civilization dies: abandoned, handcuffed by authorities who neither trusted nor cared for him, and accused o…

IndustryDGX agent

Henry Nowak died the same way a civilization dies: abandoned, handcuffed by authorities who neither trusted nor cared for him, and accused of hate crimes he did not commit. His murder is as tragic as

here's an activegraph based deep research agent that gives you full graph/trace of claims, sources, agent activity...

AgentsDGX agent

This post describes an AI research agent built on ActiveGraph that provides complete visibility into its reasoning process through detailed graphs and traces of claims, sources, and internal agent act

Here’s this week’s shipping recap 👇 — Nano Banana 2 & Nano Banana Pro are now GA and available via the Gemini Enterprise Agent Platform, Ge…

Model ReleasesDGX agent

Here’s this week’s shipping recap 👇 — Nano Banana 2 & Nano Banana Pro are now GA and available via the Gemini Enterprise Agent Platform, Gemini API, and in @GoogleAIStudio —Co-Scientist, our new multi

HERO: Learning Humanoid End-Effector Control for Visual Whole-Body Open-Vocabulary Object Grasping

SafetyDGX agent

arXiv:2602.16705v3 Announce Type: replace-cross Abstract: Visual loco-manipulation of arbitrary in-the-wild objects requires accurate end-effector (EE) control and a generalizable understanding of the

Hierarchical Mask-Enhanced Dual Reconstruction Network for Few-Shot Fine-Grained Image Classification

ResearchDGX agent

arXiv:2506.20263v2 Announce Type: replace Abstract: Few-shot fine-grained image classification (FS-FGIC) is challenging as it requires distinguishing visually similar subclasses with extremely limited

Highlights: 👉 Design-first generation for ads, posters, packaging, and product visuals 👉 Strong typography and multilingual text rendering…

ToolsDGX agent

Highlights: 👉 Design-first generation for ads, posters, packaging, and product visuals 👉 Strong typography and multilingual text rendering 👉 Precise layout and color-palette control for brand workflow

HOLO: Homography-Guided Pose Estimator Network for Fine-Grained Visual Localization on SD Maps

Model ReleasesDGX agent

arXiv:2601.02730v3 Announce Type: replace Abstract: Visual localization on standard-definition (SD) maps has emerged as a promising low-cost and scalable solution for autonomous driving. However, exis

HomeWorld: A Unified Floorplan-to-Furnished Framework for Generating Controllable, Densely Interactive Whole-Home Scenes

ResearchDGX agent

arXiv:2606.06390v1 Announce Type: new Abstract: Indoor scene generation is crucial for robot simulation and modern interior design. However, complex layouts together with scarce 3D scene data make lea

Horse Eye Blink Detection and Classification for Equine Affective State Assessment

ResearchDGX agent

arXiv:2606.05458v1 Announce Type: new Abstract: Automated detection of equine facial action units (AUs) is a promising yet under-explored avenue for pain and affective state assessment in horses. Half

Hot take: Universities charge $300,000 for a degree that teaches you skills any LLM can do for free. At some point we need to have an honest…

ApplicationsDGX agent

Hot take: Universities charge $300,000 for a degree that teaches you skills any LLM can do for free. At some point we need to have an honest conversation about whether higher education is the greatest

How a USB-connected speaker can infect a PC without ever being touched

IndustryDGX agent

A security researcher discovered that attackers can silently flash custom firmware to a Creative Sound Blaster USB speaker over Bluetooth without physical contact, and the malicious firmware can explo

How SiriusXM and Snowflake are using AI to power personalized media experiences

IndustryDGX agent

Artificial intelligence and audio audience intelligence are reshaping how media companies understand consumers and deliver personalized experiences. As organizations move beyond traditional search and

How to Stop Shipping Low-Quality RL Environments (with Examples)

TutorialsDGX agent

This article discusses best practices for ensuring high-quality reinforcement learning (RL) environments, addressing common pitfalls that lead to poor simulation design and implementation. It likely p

https://ollama.com/library/gemma4/tags

Local AiDGX agent

Gemma4 is a language model available through Ollama's model library with multiple tagged versions for different use cases and configurations. The Ollama platform enables users to run open-source large

https://www.404media.co/microsoft-wants-to-make-people-addicted-to-scout-its-new-ai-assistant-internal-documents-reveal/

ToolsDGX agent

Microsoft's internal documents reveal plans to design Scout, a new AI assistant, with features intended to increase user addiction and engagement. The strategy reportedly focuses on behavioral psychol

Human Adults and LLMs as Scientists: Who Benefits from Active Exploration?

SafetyDGX agent

arXiv:2606.06464v1 Announce Type: new Abstract: A long-standing finding in the causal learning literature is that adults struggle to identify conjunctive causal rules, where an effect requires the sim

Humans' ALMANAC: A Human Collaboration Dataset of Action-Level Mental Model Annotations for Agent Collaboration

Model ReleasesDGX agent

arXiv:2606.06388v1 Announce Type: cross Abstract: Recent advances in LLM agents have enabled complex cognitive capabilities, such as multi-step reasoning, planning, and tool use, that increasingly pos

HyperVis: Continuous Latent Visual Relational Graphs on the Lorentz Hyperboloid for Compositional Reasoning

ResearchDGX agent

arXiv:2606.06100v1 Announce Type: new Abstract: Vision-Language Models (VLMs) struggle with compositional reasoning that requires understanding inter-object relationships. A natural remedy is to injec

I absolutely agree that there really is this 10x opportunity for companies to be $40 trillion in market cap and beyond—perhaps Nvidia, Googl…

HardwareDGX agent

I absolutely agree that there really is this 10x opportunity for companies to be 40 trillion in market cap and beyond—perhaps Nvidia, Google, and beyond. Really fascinating to consider what that could

I always appreciate the opportunity to discuss @LawZero_ and our approach to honest, reliable AI. Working on the Scientist AI with my brilli…

Model ReleasesDGX agent

I always appreciate the opportunity to discuss @LawZero_ and our approach to honest, reliable AI. Working on the Scientist AI with my brilliant colleagues at LawZero has made me very confident that we

i love being (for now) bdfl for aie because i can do cheeky shit like the AGI pills we did in london and also this

ToolsDGX agent

i love being (for now) bdfl for aie because i can do cheeky shit like the AGI pills we did in london and also this @swyx @aiDotEngineer Best event in the industry. Excited to see everyone there in 3 w

I want to offer some unsolicited advice to computer vision researchers jumping into robotics. Don't focus too much on VLMs, VLAs etc. That's…

ResearchDGX agent

I want to offer some unsolicited advice to computer vision researchers jumping into robotics. Don't focus too much on VLMs, VLAs etc. That's fine, but the real action is at the sensorimotor level. Mos

IA-RAG: Interval-Algebra-Driven Temporal Reasoning for Dynamic Knowledge Retrieval

Model ReleasesDGX agent

arXiv:2606.06044v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) has shown strong effectiveness in grounding Large Language Models (LLMs) with external knowledge. However, existing

IDEAL: Leveraging Infinite and Dynamic Characterizations of Large Language Models for Query-focused Summarization

SafetyDGX agent

arXiv:2407.10486v3 Announce Type: replace-cross Abstract: Query-focused summarization (QFS) aims to produce summaries that answer particular questions of interest, enabling greater user control and pe

@ideogram_ai Ideogram 4 is now available on Together AI. Try it now: http://www.together.ai/models/ideogram-40

ToolsDGX agent

Ideogram 4, an AI image generation model, has been made available on the Together AI platform. Users can now access and experiment with Ideogram 4 through Together AI's model interface at together.ai/

If Claude is good enough for Nobel Prize winners it is good enough for you https://arxiv.org/abs/2606.03300

Model ReleasesDGX agent

This post appears to reference Claude AI's capabilities and performance, likely highlighting how the model has been used or endorsed by notable researchers or Nobel Prize winners to establish credibil

If scale was “all you need”, Elon would be hoarding LLMs, not leasing them.

HardwareDGX agent

If scale was “all you need”, Elon would be hoarding LLMs, not leasing them. Last year: people had to steal GPU’s from armored trucks. This year: SpaceX is leasing to GPUs left, right, and center, beca

Illinois Governor JB Pritzker plans to temporarily halt tax breaks for data centers from July 1, calling on state lawmakers to create a development framework (Natasha Korecki/NBC News)

IndustryDGX agent

Natasha Korecki / NBC News: Illinois Governor JB Pritzker plans to temporarily halt tax breaks for data centers from July 1, calling on state lawmakers to create a development framework — Pritzker, wh

I’m sick and tired of chat threads. Instead, for the muddiest of tasks, I make fully dynamic interfaces - some that are only used for one ta…

IndustryDGX agent

I’m sick and tired of chat threads. Instead, for the muddiest of tasks, I make fully dynamic interfaces - some that are only used for one task or one meeting. Check out the video below. I’ve got 6 rea

Imagine Before You Predict: Interleaved Latent Visual Reasoning for Video Event Prediction

ResearchDGX agent

arXiv:2606.05769v1 Announce Type: new Abstract: Video event prediction (VEP) requires models to infer unobserved future states from partial video evidence. Existing video MLLMs usually verbalize inter

Improving Answer Extraction in Context-based Question Answering Systems Using LLMs

Model ReleasesDGX agent

arXiv:2606.06197v1 Announce Type: new Abstract: Question answering (QA) systems have achieved notable progress with the advent of large language models (LLMs). However, they still face challenges in a

Improving Heart-Focused Medical Question Answering in LLMs via Variance-Aware Rubric Rewards with GRPO

Local AiDGX agent

arXiv:2606.05174v1 Announce Type: new Abstract: Large Language Models (LLMs) have shown strong promise in healthcare applications. Yet deploying general-purpose models in real-world settings remains d

In an internal message, Satya Nadella rebuked an internal memo that said Microsoft needs to 'make people addicted' to its new AI agent product called Scout (Aaron Holmes/The Information)

AgentsDGX agent

Aaron Holmes / The Information: In an internal message, Satya Nadella rebuked an internal memo that said Microsoft needs to “make people addicted” to its new AI agent product called Scout — Microsoft

In-Context Multiple Instance Learning

ApplicationsDGX agent

arXiv:2606.06458v1 Announce Type: cross Abstract: Multiple Instance Learning (MIL) addresses problems where supervision is available at the level of bags of instances and has been successfully applied

India-based Innefu Labs, which builds AI-powered software for national defense and enterprise security infrastructure, raised a $30M Series B led by Panthera (The Economic Times)

ApplicationsDGX agent

The Economic Times: India-based Innefu Labs, which builds AI-powered software for national defense and enterprise security infrastructure, raised a $30M Series B led by Panthera — The capital infusion

InfoDensity: Rewarding Information-Dense Traces for Efficient Reasoning

ResearchDGX agent

arXiv:2603.17310v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) with extended reasoning capabilities often generate verbose and redundant reasoning traces, incurring unnecessary

InfoShield: Privacy-Preserving Speech Representations for Mental Health Screening via Information-Theoretic Optimization

ResearchDGX agent

arXiv:2606.05561v1 Announce Type: new Abstract: Speech-based mental health screening offers scalable depression detection, yet clinical deployment faces a significant barrier: users' privacy concerns

Interpreting Style Representations via Style-Eliciting Prompts

ResearchDGX agent

arXiv:2606.05716v1 Announce Type: new Abstract: Style representation learning is a powerful tool for authorship analysis and modeling writing style, yet the latent nature of learned representations ma

Introducing Ideogram 4 from @ideogram_ai on Together AI, an open image model built for design with strong text rendering, layout control, an…

ApplicationsDGX agent

Introducing Ideogram 4 from @ideogram_ai on Together AI, an open image model built for design with strong text rendering, layout control, and native 2K image generation. AI natives can now use Ideogra

Inverse Design of Realizable Metasurface based Absorbers using Improved Conditioning and Diversity Enhanced Progressively Growing GANs

SafetyDGX agent

arXiv:2606.05849v1 Announce Type: cross Abstract: Metasurfaces enable precise manipulation of electromagnetic waves for applications such as beam steering, sensing, and stealth technology. However, in

Inverse Manipulation through Symbolic Planning and Residual Operator Learning

SafetyDGX agent

arXiv:2606.05248v1 Announce Type: new Abstract: Inverting a robotic task requires more than reversing symbolic state transitions or rewinding motor trajectories. In robot manipulation tasks, symbolic

← Previous
1…685686687688689…1518
Next →