AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent
84,630Total entries
1Added by human
84,629Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,603 results
8 Jun 2026

OpenGlass: Open-Source Smart Glasses for On-Device Event-Based Gesture Recognition

Model ReleasesDGX agent

arXiv:2606.07431v1 Announce Type: new Abstract: Smart eyewear enables unobtrusive, context-aware interaction through multimodal sensors and on-device intelligence, but is severely limited by power, me

OpenHalDet: A Unified Benchmark for Hallucination Detection across Diverse Generation Scenarios

Model ReleasesDGX agent

arXiv:2606.06959v1 Announce Type: cross Abstract: Hallucination detection is essential for the reliable deployment of large language models (LLMs). However, existing evaluations face two core challeng

OPTIMUS-Prime: Minimal and Sufficient Concept Explanations for Deep Vision Models

Model Releases

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2606.07180v1 Announce Type: new Abstract: The growing demand for transparency in automated decision-making has propelled eXplainable Artificial Intelligence (XAI) to the forefront of machine lea

PaperFlow: Profiling, Recommending, and Adapting Across Daily Paper Streams

Model ReleasesDGX agent

arXiv:2606.07454v1 Announce Type: cross Abstract: Scientific paper recommendation is typically evaluated as static ranking over a fixed candidate set, yet real scientific reading unfolds as a daily, l

Phun-Bench: Evaluating LLMs on Phonological Understanding in Chinese

Model ReleasesDGX agent

arXiv:2606.07300v1 Announce Type: new Abstract: Language is a vehicle for thought, intricately tied to sounds, symbols, and meaning. However, most large language model (LLM) research focuses on meanin

PhyRoGen: Synthetic Generation of Physical Robot Manipulation Puzzles Using Procedural Content Generation

Model ReleasesDGX agent

arXiv:2606.06569v1 Announce Type: new Abstract: Robot manipulation of physical puzzles is important for automatic assembly and disassembly tasks. However, to enable robots to solve physical puzzles, m

Physics-Driven Semantic Scattering Structure Understanding of Aircraft Target in SAR Images

Model ReleasesDGX agent

arXiv:2606.06847v1 Announce Type: cross Abstract: Synthetic aperture radar (SAR) has become indispensable for target interpretation owing to its all-day and all-weather observation capability. In SAR

Pipeline parallelism in llama.cpp may be wasting your VRAM

Model ReleasesDGX agent

Pipeline parallelism in llama.cpp distributes model layers across multiple GPUs, with each GPU holding a contiguous slice of layers . However, the Reddit post likely discusses inefficiencies in how pi

Principles of Concept Representation in Sentence Encoders

Model ReleasesDGX agent

arXiv:2606.06994v1 Announce Type: new Abstract: What makes a sentence encoder produce good concept representations? We approach this through the lens of representational compositionality: an encoder s

Product units in gated recurrent units improve nuclear-mass prediction

Model ReleasesDGX agent

arXiv:2606.06866v1 Announce Type: new Abstract: The prediction of masses of atomic nuclei using machine learning can complement theoretical models and advance the exploration of poorly known domains o

PromptPrint: Behavioral Biometrics Through Natural Language Prompting in LLMs

Model ReleasesDGX agent

arXiv:2606.06755v1 Announce Type: new Abstract: Authorship attribution research has traditionally focused on long-form, expressive texts; however, interactions with large language models (LLMs) are ty

PSA: Just added a few thousand chips, including B200s and B300s to our Dedicated Model Inference (http://api.together.ai/endpoints). With De…

Model ReleasesDGX agent

PSA: Just added a few thousand chips, including B200s and B300s to our Dedicated Model Inference (http://api.together.ai/endpoints). With Dedicated Model Inference, you can now on-click deploy our Bla

Quantum-Inspired Trace-Augmented Evidence Selection for Reasoning over Structured Hypothesis Spaces

Model ReleasesDGX agent

arXiv:2606.06941v1 Announce Type: new Abstract: Large language models (LLMs) now solve a wide range of expert-level exams at or above human level, yet remain brittle on specialised, evidence-intensive

RealDocBench: A Benchmark for Field-Level QA and Layout Understanding on Real-World Regulated Documents

Model ReleasesDGX agent

arXiv:2606.07401v1 Announce Type: new Abstract: Document parsing systems are increasingly deployed in high-stakes, regulated workflows such as mortgage underwriting, financial reporting, supply-chain

RECAP: Regression Evaluation for Continual Adaptation of Prompts

Model ReleasesDGX agent

arXiv:2606.06698v1 Announce Type: cross Abstract: Production agentic systems routinely face evolving constraints and must comply from the very next interaction. Scenarios like a tool-call notification

ReclAIm: A Multi-Agent Framework for Monitoring and Correcting Performance Decline in Medical Imaging AI

Model ReleasesDGX agent

arXiv:2510.17004v2 Announce Type: replace-cross Abstract: Purpose: To develop and evaluate a multi-agent framework (ReclAIm) for automated monitoring, detection, and correction of performance decline

REMEDI: A Benchmark for Retention and Unlearning Evaluation in Multi-label Clinical Disease Inference

Model ReleasesDGX agent

arXiv:2606.07141v1 Announce Type: cross Abstract: Language models trained for clinical disease inference are trained on patient data, which may include sensitive and private information, and data owne

RETROSPECT: RETROsynthesis via Sequential Prediction, and Chemically Transformed-ranking

Model ReleasesDGX agent

arXiv:2606.07181v1 Announce Type: cross Abstract: Single-step retrosynthesis needs both accurate first-ranked suggestions and candidate lists that are rich enough for downstream selection. We study th

Reversible Foundations: Training a 120B Sparse MoE through State-Preserving Scaling

Model ReleasesDGX agent

arXiv:2606.07404v1 Announce Type: new Abstract: This paper reports on training a hundred-billion-parameter sparse mixture of experts on a single eight-GPU node, end to end. LightningLM 0.1V is a recur

RhinoVLA Technical Report

Model ReleasesDGX agent

arXiv:2606.07383v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have shown strong potential for robotic manipulation, but real-time deployment on edge hardware remains challengin

RISE: Single Static Radar-based Indoor Scene Understanding

Model ReleasesDGX agent

arXiv:2511.14019v3 Announce Type: replace Abstract: Robust and privacy-preserving indoor scene understanding remains a fundamental open problem. While optical sensors such as RGB and LiDAR offer high

RPC-GS: Gaussian Splatting with native RPC Rendering for Satellite Imagery

Model ReleasesDGX agent

arXiv:2606.06690v1 Announce Type: new Abstract: We present RPC-GS, the first Gaussian Splatting framework for satellite imagery that operates natively with Rational Polynomial Camera (RPC) models. The

Rubrics are even more flexible than /goal You can define a custom subagent for the grading, including custom tools, prompt, and iteration li…

Model ReleasesDGX agent

Rubrics are even more flexible than /goal You can define a custom subagent for the grading, including custom tools, prompt, and iteration limits. Try it out and let us know what you think! we just shi

ScenicRules: An Autonomous Driving Benchmark with Multi-Objective Specifications and Abstract Scenarios

Model ReleasesDGX agent

arXiv:2602.16073v2 Announce Type: replace-cross Abstract: Developing autonomous driving systems for complex traffic environments requires balancing multiple objectives, such as avoiding collisions, ob

SEAM: Shortcut-Aware Real-Time Detection of Scripted vs. Spontaneous Speech for Interview Guardrails

Model ReleasesDGX agent

arXiv:2606.06837v1 Announce Type: cross Abstract: Scripted vs spontaneous speech detection is appealing for interview guardrails, but benchmark performance can be inflated by shortcuts tied to corpus

Seeing a number of benchmarks showing Opus is the best model for long-running work. Five tips for running Opus autonomously for hours/days: …

Model ReleasesDGX agent

Seeing a number of benchmarks showing Opus is the best model for long-running work. Five tips for running Opus autonomously for hours/days: 1. Use auto mode for permissions, so Claude doesn’t ask for

Seeing Without Exposing: Adaptive Privacy Control for Open-World, Context-Hungry MLLMs

Model ReleasesDGX agent

arXiv:2606.07175v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have raised new privacy challenges. On the data side, user-provided inputs often include unpredictable sensitiv

ShallowBench: Benchmarking Generative Drug Design Models on Shallow-Pocket Targets

Model ReleasesDGX agent

arXiv:2606.06717v1 Announce Type: cross Abstract: While generative AI models have demonstrated remarkable success in structure-based drug design, they predominantly rely on deep binding pockets and st

SigmaScale: LLM Compression with SVD-based Low-Rank Decomposition and Learned Scaling Matrices

Model ReleasesDGX agent

arXiv:2606.07098v1 Announce Type: new Abstract: We present SigmaScale, a method for learning auxiliary scaling matrices S to aid truncated Singular Value Decomposition (SVD) based Large Language Model

Siri AI at WWDC 2026

Model ReleasesDGX agent

Given how badly burned anyone who took Apple's 2024 WWDC Apple Intelligence announcements at face value was, I'm holding to a strict 'I'll believe it when I see it' policy for everything they announce

Skill-3D: Evolving Scene-Aware Skills for Agentic 3D Spatial Reasoning

Model ReleasesDGX agent

arXiv:2606.07436v1 Announce Type: new Abstract: This paper explores agentic 3D spatial understanding, i.e., MLLM agents performing 3D reasoning through tool use. Existing methods often misuse tools an

So excited to be opening up OpenEnv to the whole community. It will now be owned by @huggingface , Meta-PyTorch, @reflection_ai , @UnslothAI…

Model ReleasesDGX agent

So excited to be opening up OpenEnv to the whole community. It will now be owned by @huggingface , Meta-PyTorch, @reflection_ai , @UnslothAI , @modal, @PrimeIntellect , @NVIDIAAI , @mercor_ai , and @f

Sparse Subspace-to-Expert Sharing for Task-Agnostic Continual Learning

Model ReleasesDGX agent

arXiv:2606.07500v1 Announce Type: cross Abstract: Continual learning in Large Language Models (LLMs) is hindered by the plasticity-stability dilemma, where acquiring new capabilities often leads to ca

Spatial-Temporal Decoupled Adapter for Micro-gesture Online Recognition

Model ReleasesDGX agent

arXiv:2606.07355v1 Announce Type: new Abstract: Micro-gesture online recognition aims to temporally localize and classify subtle gestures in untrimmed videos. Owing to their extremely short duration,

Spline Policy: A Structured Representation for Robot Policies

Model ReleasesDGX agent

arXiv:2606.07386v1 Announce Type: new Abstract: Modern imitation-learning policies for robot manipulation often represent actions as fixed-resolution action chunks, which are simple and effective but

STREAM: Stochastic Riemannian Flow Matching with Anisotropic Decoder for Digital Histopathology Image Generation

Model ReleasesDGX agent

arXiv:2606.07036v1 Announce Type: cross Abstract: Synthetic histopathology image generation addresses critical challenges in computational pathology, including patient privacy and the growing need for

Stream3D-VLM: Online 3D Spatial Understanding with Incremental Geometry Priors

Model ReleasesDGX agent

arXiv:2606.06891v1 Announce Type: new Abstract: Despite advances in 3D scene understanding, existing 3D Large Multimodal Models operate in offline settings, requiring complete scene observations or pr

Style or Content? Evaluating Style Classifiers with Controlled Content Overlap

Model ReleasesDGX agent

arXiv:2606.07103v1 Announce Type: new Abstract: Style classifiers can use content cues that correlate with style labels in naturally collected data, yet we lack a systematic way to measure this relian

Superintelligent Retrieval Agent: The Next Frontier of Agentic Retrieval

Model ReleasesDGX agent

arXiv:2605.06647v2 Announce Type: replace-cross Abstract: Retrieval-augmented agents are increasingly the interface to large knowledge bases, yet most treat retrieval as a black box: they issue explor

Supervision versus Demonstration-Based In-Context Learning for Multiword Expression Classification

Model ReleasesDGX agent

arXiv:2606.07479v1 Announce Type: cross Abstract: Turkish idiomatic light verb constructions (LVCs) are challenging for multiword expression processing because they often share the same surface form a

SVHighlights: Towards Extremely Long Sport Video Highlight Detection

Model ReleasesDGX agent

arXiv:2606.06926v1 Announce Type: new Abstract: While highlight detection for long-form videos is of great practical importance, most existing methods remain limited to short-form content, largely due

SW-A^2-Bench: Benchmarking Autonomous Software Agent Generation for Agentic Web

Model ReleasesDGX agent

arXiv:2604.04226v2 Announce Type: replace-cross Abstract: The Agentic Web is emerging as a paradigm in which autonomous software agents interact with online resources and with each other to accomplish

SWE-Explore: Benchmarking How Coding Agents Explore Repositories

Model ReleasesDGX agent

arXiv:2606.07297v1 Announce Type: cross Abstract: Repository-level coding benchmarks such as SWE-bench have driven a rapid surge in the capabilities of coding agents. Yet they usually treat coding tas

TALAN: Task-Aligned Latent Adaptation Networks for Targeted Post-Training of Large Language Models

Model ReleasesDGX agent

arXiv:2606.06902v1 Announce Type: new Abstract: Targeted post-training aims to improve reasoning, math, and code without degrading strengths. Low-rank adapters are efficient but task-global; activatio

Test-Time Trajectory Optimization for Autonomous Driving

Model ReleasesDGX agent

arXiv:2606.07170v1 Announce Type: new Abstract: End-to-end planners for autonomous driving typically generate a set of candidate trajectories, score each one, and return the highest-scoring candidate.

TEVI: Text-Conditioned Editing of Visual Representations via Sparse Autoencoders for Improved Vision-Language Alignment

Model ReleasesDGX agent

arXiv:2606.07451v1 Announce Type: cross Abstract: Vision-language models such as CLIP are highly useful for diverse tasks due to their shared image-text embedding space. Despite this, the image and te

Textual Supervision Enhances Geospatial Representations in Vision-Language Models

Model ReleasesDGX agent

arXiv:2606.07172v1 Announce Type: cross Abstract: Geospatial understanding is a critical yet underexplored dimension in the development of machine learning systems for tasks such as image geolocation

The Fine-Tuning Trap: Evaluating Negative Transfer and the Role of PEFT in Sub-1B Mathematical Reasoning

Model ReleasesDGX agent

arXiv:2606.06920v1 Announce Type: cross Abstract: Deploying Small Language Models (SLMs) on edge devices requires efficient fine-tuning strategies that adapt models to new tasks without degrading thei

The Geometry of Representational Failures in Vision Language Models

Model ReleasesDGX agent

arXiv:2602.07025v2 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) exhibit puzzling failures in multi-object visual tasks, such as hallucinating non-existent elements or failing t

The highest leverage work in AI right now is some of the most boring. (Well boring to others, I kind of love the pain of problem-solving.) E…

Model ReleasesDGX agent

The highest leverage work in AI right now is some of the most boring. (Well boring to others, I kind of love the pain of problem-solving.) Everyone and their boss wants to build the cool AI agent... t

The north stars we're working towards at OpenAI all center around the mission: ensure AGI benefits all of humanity. AI should expand human a…

Model ReleasesDGX agent

The north stars we're working towards at OpenAI all center around the mission: ensure AGI benefits all of humanity. AI should expand human agency, not make people less consequential to the future. htt

The Piggyback Hypothesis of Generalization: Explaining and Mitigating Emergent Misalignment

Model ReleasesDGX agent

arXiv:2606.06667v1 Announce Type: new Abstract: The mechanisms behind LLMs' broad over-generalization beyond training examples remain unclear. Emergent misalignment (EM) offers a striking case study:

The Post-GCN Decade Revisited: Curvature-Stratified Evaluation of Relational Learning

Model ReleasesDGX agent

arXiv:2606.06397v2 Announce Type: replace Abstract: Current evaluation practices in relational learning rely heavily on flat leaderboards that average performance across heterogeneous datasets, implic

There’s so much demand for a good small model, look at top downloaded qwen models All < 9b

Model ReleasesDGX agent

The post highlights strong market demand for efficient small language models under 9 billion parameters, citing Qwen's top-downloaded models as evidence of user preference for compact, resource-effici

Think Fast: Estimating No-CoT Task-Completion Time Horizons of Frontier AI Models

Model ReleasesDGX agent

arXiv:2606.07157v1 Announce Type: new Abstract: Many efforts to ensure frontier AI models are safe rely on monitoring their chain-of-thought (CoT) reasoning. If models become able to perform sufficien

Think Like a Pilot: Fine-Grained Long-Horizon UAV Navigation

Model ReleasesDGX agent

arXiv:2606.06836v1 Announce Type: cross Abstract: Language-guided UAV agents must execute long-horizon semantic instructions while producing smooth, physically feasible continuous flight commands, yet

ThinkBooster: A Unified Framework for Seamless Test-Time Scaling of LLM Reasoning

Model ReleasesDGX agent

arXiv:2606.06915v1 Announce Type: cross Abstract: Test-time compute (TTC) scaling has emerged as a powerful paradigm for improving large language model (LLM) reasoning by allocating additional compute

This is super big I think this is the first useful speculative decoding method deployed on a big quasi frontier model Massive unlock @fi5662…

Model ReleasesDGX agent

This is super big I think this is the first useful speculative decoding method deployed on a big quasi frontier model Massive unlock @fi56622380 🚀 1,000+ TOKENS/S ON A 1T MODEL! 🚀 We are thrilled to r

TokaMind: A Multi-Modal Transformer Foundation Model for Tokamak Plasma Dynamics

Model ReleasesDGX agent

arXiv:2602.15084v2 Announce Type: replace-cross Abstract: We present TokaMind, to our knowledge the first open-source foundation model for tokamak plasma dynamics, based on a Multi-Modal Transformer (

Tokyo this week

Model ReleasesDGX agent

This post likely provides weekly updates, events, or recommendations for activities and happenings in Tokyo for the current week. The content probably covers entertainment, dining, cultural events, or

← Previous
1…162163164165166…377
Next →