AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
59,929 results
31 Jul 2026

THGFM: Dual-Branch Temporal Heterogeneous Graph Fusion Model

Model ReleasesDGX agent

arXiv:2607.27303v1 Announce Type: cross Abstract: Temporal heterogeneous graphs offer a natural abstraction for dynamic relational systems in which diverse node and relation types co-exist and evolve

Using Large Language Models for Idea Generation in Innovation

Model ReleasesDGX agent

arXiv:2607.27553v1 Announce Type: cross Abstract: This research evaluates the efficacy of large language models (LLMs) in generating new product ideas. To do so, we compare three pools of ideas for ne

What Makes Graph Unified? Principles and Generative Sliding-Window Transformer for Graph Foundation Models

TutorialsDGX agent

arXiv:2607.27966v1 Announce Type: new Abstract: Graph Foundation Models (GFMs) have recently emerged as a promising paradigm for general-purpose graph learning, aiming to learn reusable knowledge that

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
30 Jul 2026

ActSWM: Action-Sensitive World Models for Long-Horizon Planning in Open-World Games

Local AiDGX agent

arXiv:2607.26712v1 Announce Type: new Abstract: Latent world models support efficient model-predictive control by optimizing future control sequences in latent space and replanning in a receding-horiz

ARC-Encoder: learning compressed text representations for large language models

Model ReleasesDGX agent

arXiv:2510.20535v2 Announce Type: replace Abstract: Recent techniques such as retrieval-augmented generation or chain-of-thought reasoning have led to longer contexts and increased inference costs. Co

Archetypes or ability? Clustering for modelling student mathematical competence

Model ReleasesDGX agent

arXiv:2607.26063v1 Announce Type: cross Abstract: Personalised learning systems often assume that mathematical ability is combined of discrete abilities, acquired sequentially and dependent upon first

From Conceptual Hydrologic Models to Conceptually Interpretable Neural Networks: A Snow-Water Mass-Conserving-Perceptron Framework for Discovering Catchment-Scale Precipitation-Storage-Runoff Representations

AgentsDGX agent

arXiv:2607.26492v1 Announce Type: new Abstract: The Mass-Conserving Perceptron (MCP) establishes a modeling paradigm in which conceptual hydrologic models can be reformulated as physically constrained

HumanCLAW: Can Vision-Language Models Act Through a Body?

Model ReleasesDGX agent

arXiv:2607.27180v1 Announce Type: new Abstract: Evaluating whether a vision-language model (VLM) can act through a physical body is challenging. The outcome of an action couples the VLM's decision wit

OptimismBench: Forecasting Bias and the Alignment Effect in Language Model Judgment

SafetyDGX agent

arXiv:2607.26981v1 Announce Type: new Abstract: Large language models are increasingly used as decision aids whose probability judgments shape downstream choices. Whether those judgments carry a syste

Relation Geometry in Semantic Space of Language Models

ResearchDGX agent

arXiv:2607.26762v1 Announce Type: new Abstract: When it comes to generating vector representations of words, current language models are achieving high-quality results. However, what is not known is t

Scientific Knowledge Discovery in the Age of Large Language Models

ResearchDGX agent

arXiv:2607.26670v1 Announce Type: cross Abstract: The rapid growth of scholarly literature has made identifying relevant publications increasingly difficult, and conventional search systems still depe

Thinking Under Uncertainty: Evidence Use and Information-Seeking in Language Models

SafetyDGX agent

arXiv:2607.26845v1 Announce Type: new Abstract: Inference-time thinking improves the performance of large language models, but aggregate outcomes do not reveal whether models use available evidence mo

We are committed to pushing the model frontier across cost efficiency, capability, and speed. Starting today, we are reducing prices for GPT…

Model ReleasesDGX agent

We are committed to pushing the model frontier across cost efficiency, capability, and speed. Starting today, we are reducing prices for GPT-5.6 Luna by 80% and GPT-5.6 Terra by 20% , and offering a f

29 Jul 2026

CogArena: A Multimethod Evaluation of Cognitive Ability Structure in Large Language Models

Model ReleasesDGX agent

arXiv:2607.24999v1 Announce Type: cross Abstract: LLM cognitive scores are increasingly summarized as per-ability profiles whose dimensions should converge across tasks, respond selectively to matched

CoTinyVLA: Chain-of-Thought Distillation for a Sub-Billion-Parameter Vision-Language-Action Model

Model ReleasesDGX agent

arXiv:2607.25487v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models translate natural-language commands into robot action sequences, but leading systems on the LIBERO-Plus robustness b

From CUDA to MLX: How K-Search Brings Decades of Kernel Expertise to Apple Silicon

Model ReleasesDGX agent

Figure 1: CUDA-to-MLX optimization translation map. CUDA optimization knowledge can be translated into architecture-native MLX strategies rather than copied instruction-for-instruction. We face a new

How Small Can You Go? A Controlled Study of LoRA Rank, Target Modules, and Quantization Trade-offs for Text-to-SQL on a 60M-Parameter Model

Model ReleasesDGX agent

arXiv:2607.25583v1 Announce Type: new Abstract: Parameter-efficient fine-tuning (PEFT) and low-bit quantization are now standard tools for adapting language models under tight compute budgets, yet the

Mage-VL: An Efficient Codec-Native Streaming Multimodal Foundation Model

Model ReleasesDGX agent

arXiv:2607.24904v1 Announce Type: cross Abstract: Standard vision-language models (VLMs) suffer from Moravec's paradox: they excel at complex offline visual reasoning but struggle with simple streamin

Raven: High-Recall Sequence Modeling with Sparse Memory Routing

ResearchDGX agent

arXiv:2607.25357v1 Announce Type: cross Abstract: Long-context recall in linear-time sequence models highlights a tradeoff in how they write to memory. State-based linear models, such as state-space m

Similar Models Learn Differently: Final-Window Pretraining Shapes Post-Training Beyond SFT

Model ReleasesDGX agent

arXiv:2607.25063v1 Announce Type: new Abstract: Developers judge a model checkpoint by how it behaves. After supervised fine-tuning (SFT), two checkpoints that perform about the same across relevant b

Towards Understanding the Cognitive Habits of Large Reasoning Models

Model ReleasesDGX agent

arXiv:2506.21571v3 Announce Type: replace-cross Abstract: Large Reasoning Models (LRMs), which autonomously produce a reasoning Chain of Thought (CoT) before producing final responses, offer a promisi

28 Jul 2026

Action from Adjacent Set in Physical Space Outperforms the Best Prediction in World Models

ResearchDGX agent

arXiv:2607.23602v1 Announce Type: cross Abstract: Controllers based on sampling and latent world models assign a predicted terminal cost to each candidate action sequence, choose the minimum, execute

Autoregressive One-Step Generative Modeling for Dynamical System Forecasting

ResearchDGX agent

arXiv:2605.05540v2 Announce Type: replace Abstract: Fast surrogate modeling for high-dimensional physical dynamics requires more than low short-term error: useful models must roll out efficiently whil

CONSISTRE: A Unified Consistency-Aware Framework for Document-Level Relation Extraction with Large Language Models

Local AiDGX agent

arXiv:2607.24312v1 Announce Type: new Abstract: Document-level relation extraction (DocRE) aims to extract relations among multiple entities across extended contexts while maintaining consistency acro

ERUnderstand: Evaluating Vision-Language Models on Structured ER Diagrams

Model ReleasesDGX agent

arXiv:2607.24707v1 Announce Type: new Abstract: Entity-Relationship Diagrams (ERDs) are central to conceptual database design, yet they are typically available only as rendered images rather than mach

Frustratingly Simple Black-Box Adaptation of Language Models via Logit Bias

SafetyDGX agent

arXiv:2607.22837v1 Announce Type: cross Abstract: Many organizations aim to adapt language models for internal use, both to improve performance on domain-specific tasks and to address privacy concerns

Hierarchical Reinforcement Learning with Optimal Level Synchronization Based on Flow-Based Deep Generative Model

Model ReleasesDGX agent

arXiv:2107.08183v2 Announce Type: replace Abstract: High-dimensional state and action spaces combined with sparse reward structures in reinforcement learning (RL) environments typically require advanc

Like a bilingual baby: The advantage of visually grounding a bilingual language model

TutorialsDGX agent

arXiv:2210.05487v3 Announce Type: replace Abstract: Unlike most neural language models, humans learn language in a rich, multi-sensory and, often, multi-lingual environment. Current language models ty

Moving-Horizon Estimation and Nonlinear Model Predictive Control of Cable-Driven Soft Manipulators

ResearchDGX agent

arXiv:2607.24029v1 Announce Type: new Abstract: Precise control of soft manipulators remains challenging due to the difficulty of developing accurate yet computationally tractable models for model-bas

Real-Time Human-Centric World Modeling for Upper-Body Human-Object Interaction

Local AiDGX agent

arXiv:2607.23517v1 Announce Type: new Abstract: We present a real-time human-centric world model for upper-body interactive generation, aiming to synthesize coherent local world dynamics centered on a

Self-Boosting Vision-Language Models with Noisy Student On-Policy Self-Distillation

SafetyDGX agent

arXiv:2607.23125v1 Announce Type: new Abstract: Post-training enables vision-language models (VLMs) to understand human instructions and perform various downstream tasks. Current post-training methods

SQBench: A Benchmark for Evaluating Task Delivery by Language-Model Agents in Production-Oriented Workflows

Model ReleasesDGX agent

arXiv:2607.23123v1 Announce Type: new Abstract: Existing evaluations of large language models cover knowledge, reasoning, coding, and tool use, but they rarely treat a verifiable deliverable produced

The Agentic AI Foundation updates MCP with a fully stateless architecture, a hardened authentication model, a formal 12-month deprecation policy, and more (Michael Nuñez/VentureBeat)

SafetyDGX agent

Michael Nuñez / VentureBeat: The Agentic AI Foundation updates MCP with a fully stateless architecture, a hardened authentication model, a formal 12-month deprecation policy, and more — The Model Cont

The apps you use swap AI models constantly. Doing that without downtime or a bad rollout reaching users is still often manual. We rebuilt th…

ToolsDGX agent

The apps you use swap AI models constantly. Doing that without downtime or a bad rollout reaching users is still often manual. We rebuilt that workflow based on what we’ve learned serving more than 40

The balance between compactness and forecast accuracy of data-driven latent-space reduced-order models in controlled wake flows

TutorialsDGX agent

arXiv:2607.24569v1 Announce Type: cross Abstract: Model-based active flow control requires predictive models that are accurate, stable, and fast enough for real-time optimisation. In controlled wake f

The Kimi K3 architecture figure for yesterday's big open-weight model release, along with some observations and thoughts. 1. Yes, it looks r…

Model ReleasesDGX agent

The Kimi K3 architecture figure for yesterday's big open-weight model release, along with some observations and thoughts. 1. Yes, it looks relatively complicated, but it's essentially a scaled-up prod

TriSP: Tri-Signal Structured Pruning for Large Language Models

Model ReleasesDGX agent

arXiv:2607.22587v1 Announce Type: new Abstract: Large language models (LLMs) achieve strong performance across diverse tasks but their deployment is constrained by the memory and compute cost of their

Very interesting paper on LLM reasoning. They find that frontier models can exhibit invisible reasoning by leveraging semantically irrelevan…

TutorialsDGX agent

Very interesting paper on LLM reasoning. They find that frontier models can exhibit invisible reasoning by leveraging semantically irrelevant filler tokens. In other words, invisible reasoning can ser

Which Models Perform Better in Inheritance Reasoning?

Model ReleasesDGX agent

arXiv:2606.13751v4 Announce Type: replace Abstract: This paper presents the participation of team PSL in the QIAS 2026 Shared Task on Arabic Islamic inheritance reasoning. The task evaluates the abili

27 Jul 2026

A Consensus-Based Framework for Relative Preference Evaluation of Large Language Models

SafetyDGX agent

arXiv:2607.21632v1 Announce Type: new Abstract: Traditional benchmarks for LLMs primarily rely on static datasets and objective scoring metrics, which often fail to capture differences in response qua

b10144

Model ReleasesDGX agent

server + ui: fix stream routes for model names containing a slash (#26137) server + ui: refactor resumable stream routes to query string conv_id The conversation id can embed a model name containing s

Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models

Model ReleasesDGX agent

arXiv:2607.21617v1 Announce Type: cross Abstract: Vision Language Models (VLMs) are increasingly used in place of traditional OCR pipelines for document understanding. In this paper, we show they do n

Kimi K3 is now available for inference & training in @FireworksAI_HQ. Crazy how easy they make it to tune frontier open models like K3 using…

Model ReleasesDGX agent

Kimi K3 is now available for inference & training in @FireworksAI_HQ. Crazy how easy they make it to tune frontier open models like K3 using LoRA adapters. Best time to figure out how to own your inte

Meta-Learning Approaches for Speaker-Dependent Voice Fatigue Models

ApplicationsDGX agent

arXiv:2505.23378v3 Announce Type: replace Abstract: Speaker-dependent modelling can substantially improve performance in speech-based health monitoring applications. While mixed-effect models are comm

Physically Constrained Federated Additive Models for O-RAN SLA-Risk Prediction

Local AiDGX agent

arXiv:2607.21665v1 Announce Type: new Abstract: Proactive service assurance in O-RAN requires predicting per-slice SLA violations before they occur. The prediction model must be auditable by operators

Quantum Spectral Model: Data Reuploading with Input-Conditioned Frequency Support

Local AiDGX agent

arXiv:2607.22516v1 Announce Type: cross Abstract: A central design principle in modern machine learning and artificial intelligence is to align a model's inductive bias with the structure of its input

24 Jul 2026

A joint preliminary evaluation by the UK's AISI and the US' CAISI finds Kimi K3 trails leading US frontier closed weight models on cyber capability (AI Security Institute)

IndustryDGX agent

AI Security Institute: A joint preliminary evaluation by the UK's AISI and the US' CAISI finds Kimi K3 trails leading US frontier closed weight models on cyber capability — The UK Artificial Intellige

Black box behavioural modelling: Predicting human activity schedules with a deep conditional generative approach

ResearchDGX agent

arXiv:2512.04223v2 Announce Type: replace Abstract: Modelling the complexity and diversity of human activity scheduling behaviour is inherently challenging. We demonstrate ActVAE, a deep conditional-g

Introducing Claude Opus 5 on AWS: Anthropic’s most capable Opus model

Model ReleasesDGX agent

This post covers Opus 5’s improvements and practical guidance for AI engineers integrating the model into agentic systems and production inference workloads on Amazon Bedrock. See the documentation fo

What I learned using Ollama on a real Paperless archive: model choice was not the main problem

Local AiDGX agent

I maintain Tagvico, an open-source companion for Paperless-ngx. I added Ollama because document text is exactly the kind of data many people do not want to send to a hosted model. The surprising failu

Wisdom of LLM Crowds: Aggregation and Contamination in Language Model Ensembles

Local AiDGX agent

arXiv:2607.18269v2 Announce Type: replace Abstract: The wisdom of crowds -- the finding that aggregating judgments across individuals often outperforms the best individual -- has been extensively stud

23 Jul 2026

Information Discernment in Large Language Models

Model ReleasesDGX agent

arXiv:2607.19355v1 Announce Type: new Abstract: LLMs are increasingly used with external knowledge sources like the internet. Do they weigh information appropriately -- updating more for reliable sour

MV-Bench: Benchmarking Multimodal Large Language Models for Coordinated Multi-View Interface Construction

Model ReleasesDGX agent

arXiv:2607.19910v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) are increasingly expected to automate visualization development by generating code directly from visual designs

PathAgentBench: Benchmarking Evidence-Seeking Vision-Language Models on Whole-Slide Pathology Image

Model ReleasesDGX agent

arXiv:2607.19261v1 Announce Type: new Abstract: Whole-slide image (WSI) diagnosis requires identifying diagnostically relevant regions, examining them across magnifications, and integrating multi-scal

Read or Ignore? A Unified Benchmark for Typographic-Attack Robustness and Text Recognition in Vision-Language Models

Model ReleasesDGX agent

arXiv:2512.11899v2 Announce Type: replace Abstract: Large vision-language models (LVLMs) are vulnerable to typographic attacks, where misleading text inserted into an image can override visual underst

Statistically Grounded Sparse-Feature Interventions for Activation-Space Control in Large Language Models

Model ReleasesDGX agent

arXiv:2607.19364v1 Announce Type: new Abstract: Activation steering offers a lightweight alternative to fine-tuning for behavioral control of large language models, but SAE-based steering methods ofte

Strong Gravitational Lensing Posterior Sampling in Pixel-Space Using Diffusion Models and Recurrent Inference Machines

ResearchDGX agent

arXiv:2607.19459v1 Announce Type: cross Abstract: Modeling galaxy-galaxy strong gravitational lenses to infer the brightness of the source galaxy and the mass distribution of the foreground galaxy is

WorldPack: Dynamic Frame Compression for Long-context Video World Modeling

Model ReleasesDGX agent

arXiv:2512.02473v2 Announce Type: replace-cross Abstract: Video world models have attracted significant attention for their ability to produce high-fidelity future visual observations conditioned on p

21 Jul 2026

Incredibly proud of the Sakana AI team. We have developed an orchestration model right here out of Japan that achieves state-of-the-art perf…

Model ReleasesDGX agent

Incredibly proud of the Sakana AI team. We have developed an orchestration model right here out of Japan that achieves state-of-the-art performance on real-world cybersecurity benchmarks! 🎌 Introducin

20 Jul 2026

SSI should open-source an opus-tier model. 1. it nearly closes the us/china gap on the os frontier 2. they don't want to waste compute alloc…

SafetyDGX agent

SSI should open-source an opus-tier model. 1. it nearly closes the us/china gap on the os frontier 2. they don't want to waste compute allocation on inference, releasing os will enable continued full

← Previous
1…6061626364…999
Next →