As API costs for general-purpose LLMs rise, relying solely on off-the-shelf models can quickly undermine both cost control and system reliability. In this session, we share how Nearmap moved beyond API dependency by fine-tuning and distilling domain-specific models on AWS to analyze 300 million building permits for roof modifications. Well discuss our approach to generating and structuring training data, distilling large models into smaller, production-ready alternatives, evaluating trade-offs across model architectures, and making data-driven accuracy-versus-cost decisions before deployment. Attendees will leave with concrete patterns for shipping efficient, specialized models into production.
What this session is about
Playbook
Editorial commentary · what to actually do about this on Monday
Independent editorial perspective — not an official AWS or speaker statement. Designed for executives evaluating what to brief their teams on next.
Live updates related to this session LIVE
Sourced via Parallel AI Monitor — continuous web watch on 21 topical streams. Updated .
- coursiv.io high confidence General tech / AI / startup news
Gemini 3.6 Flash: Price, Benchmarks, API & Flash-Lite
Google released Gemini 3.6 Flash as a generally available, production-ready model on July 21, 2026.
- biopharmadive.com high confidence General tech / AI / startup news
Dimension restocks with $800M to capitalize on AI's 'rapid ...
Google released Gemini 3.6 Flash as a generally available, production-ready model on July 21, 2026.
- epoch.ai General tech / AI / startup news
Data on AI Models
Notion introduced a redesigned AI model picker that presents a shortlist of models for demanding tasks and adds scorecards comparing each model’s speed, intelligence, and cost. This materially changes how users select and evaluate models within Notion.
- agentmarketcap.ai high confidence Agent frameworks (LangGraph, CrewAI, AutoGen)
AutoGen in Maintenance Mode: Microsoft Agent Framework 1.0 ...
Microsoft AutoGen transitioned to maintenance mode and is now community-managed, replaced by the official production-ready Microsoft Agent Framework (MAF) 1.0.
- openai.com General tech / AI / startup news
OpenAI appoints Dali Rajic as Chief Revenue Officer
OpenAI previewed Ultrafast, a new API service tier that runs GPT-5.6 Sol at up to 14 times the speed of Standard processing. The tier is designed to provide a new high-speed class for frontier-model inference and is powered by Cerebras infrastructure, materially expanding the per
External links matched to this session via topic relevance. The KB does not endorse third-party content; verify before citing.