The US DOE Genesis Mission "AI for Science Fellowship" (9–12 months embedded at INL/Brookhaven/PPPL; $200K prorated) has its application deadline July 31, 2026 — an in-window action date. CBAI's nine-week Summer Research Fellowship in AI Sa
Learn
Learn AI
Research artifacts, explainers, resources, benchmark literacy, and practical learning coverage.
12 sourced postsGoogle DeepMind CEO Demis Hassabis proposed an industry-funded, federally supervised body to evaluate frontier models pre-release (voluntary first, potentially mandatory certification later), plus an international watchdog — the same week C
Two practitioner references frame July's headline agentic metric: qaskills.sh's guide explains Terminal-Bench's end-state verification design (Docker sandbox, pytest-style checks of machine final state, not transcripts; Stanford × Laude Ins
OpenAI's July 20 post on safety/alignment for long-horizon models (the Erdős-model disclosure, item R8 below) lays out defense-in-depth, trajectory-level monitoring, and incident-driven red-teaming — already being used as teaching material
A cluster of practitioner guides prepares teams for the July 27 weights drop: kimik3.io's weights tracker does "the hardware arithmetic" (2.8T params ≈1.4TB at 4-bit → ≥8×192GB accelerators floor, before KV cache for 1M context); digitalapp
geotoolbox.ai published two widely cited explainers: "What Is Kimi K3?" (Jul 18) distinguishing vendor claims from independent checks, and "Open Weights vs Open Source: The Real Difference" (Jul 19), using K3's announcement-to-weights gap a
joinleland.com published a balanced primer on mechanistic interpretability's capabilities and limits (partial coverage, scale problem, cross-model generalization, false-security risk), timed to Anthropic's global-workspace/J-space result (w
Apple released the iOS 27 public beta Jul 13–14, opening the Gemini-powered Siri AI rebuild (personal context, on-screen awareness, cross-app actions, standalone Siri app, Dynamic Island invocation) beyond developers for the first time; Tec
knolli.ai's comparison of 10 agentic frameworks (LlamaIndex, LangGraph, Google ADK et al.) and Morphisec's prompt-injection/model-poisoning/supply-chain explainer circulated July 13; same day Google added agent workflows to open-source Genk
Reuters exclusive (Jul 9, memo reviewed): Iris (Broadcom-designed, TSMC-built, 4th-gen MTIA) enters production September after six-week bug validation with no major issues; 7 GW deployed 2026 doubling to 14 GW in 2027; up to $145B 2026 AI i
AIMultiple published downloadable LLM latency benchmark data (1.3K data points, CSV+README) across use cases; zylos.ai's evaluation guide maps saturated benchmarks (MMLU, GSM8K, HumanEval) vs. current differentiators (GPQA, SWE-bench Pro, M
MarkTechPost published a runnable tutorial that builds a miniature omnimodal Mixture-of-Transformers world model mirroring Cosmos-3's design (shared cross-modal attention + modality-specific expert routing for text/vision/action), with synt