
Meta Adds NTS Authenticated Time Service at nts.meta.com
Meta's public time service now supports NTS (Network Time Security, RFC 8915) at nts.meta.com. Packets are authenticated so devices can verify time came from Meta and was not modified in transit. Meta says its NTS servers hold no per-client state, with cookie keys derived rather than stored or replicated, and it has open sourced the implementation.
Sources and evidence
Summary last validated Oct 6, 2026
Reader actions
Report an issue
Use this for an incorrect summary, wrong source, duplicate story, or wrong category. Submissions are private and do not change the story automatically.
Related coverage
- infrastructure
Ai2 replaces priority GPU scheduler with budget-based fair-share system
Ai2's AI Infrastructure team replaced its priority-based GPU scheduler with a system using GPU time budgets, hierarchical fair-share allocation, and time-slicing contracts. The institute manages thousands of NVIDIA H100, B200, and B300 GPUs across 88- to 1024-GPU clusters for about 150 researchers, with demand running 2-3x available capacity. The change moved GPU allocation debates into a transparent administrative budgeting process.
- infrastructure
Ai2 details GPU scheduler using time budgets and fair-share allocation
Ai2 published a blog post explaining its new GPU scheduler for its clusters. The scheduler combines time budgets, fair-share allocation, and time-slicing to prioritize high-impact research, shorten queue waits, and keep GPUs busy.
- infrastructure
Cohere Publishes Guidance on Shared vs Dedicated Inference for Embed and Rerank
Cohere published a blog post explaining how to choose between shared, consumption-based inference and dedicated, provisioned inference for Embed and Rerank workloads. It argues request shape matters more than request volume, that traffic patterns determine cost efficiency, and that break-even depends on request size, token volume, and reranking candidate count rather than infrastructure pricing alone.
- infrastructure
PyTorch details session-aware agentic inference with NVIDIA Dynamo
PyTorch published a blog post describing session-aware agentic inference with NVIDIA Dynamo. It explains that agentic workloads differ from single-turn chat: an agent session can include a large initial prefill, repeated model calls, and parallel subagents, changing the traffic an inference server handles.
- infrastructure
AWS details multi-team GPU sharing on SageMaker HyperPod
AWS published a reference architecture for sharing one Amazon SageMaker HyperPod EKS cluster across multiple teams. It uses AWS IAM Identity Center for authentication, per-team SageMaker Domains and Kubernetes namespaces for isolation, HyperPod Task Governance for fairness, and namespace-level cost allocation for chargeback.