Hi there, this is your daily βοΈ Devshot.
In today's Devshot:
π Meta doubled the efficiency of its ads AI model
π Cloudflare Workers now support TCP
β‘ Vercel eve builds Next.js for AI agents
π Kubernetes Gateway API adds stable TCP and UDP routing
π NuGet API keys now expire in 30 days
Plus: π 13 other news you might like, π§° 6 tools, and π 5 papers.
π Meta doubled the efficiency of its ads AI model LINK
π Cloudflare Workers now support TCP LINK
β‘ Vercel eve builds Next.js for AI agents LINK
π Kubernetes Gateway API adds stable TCP and UDP routing LINK
π NuGet API keys now expire in 30 days LINK
Other news you might like
- How to implement JWT authentication in NestJSLINK
- Azure and Community Guidelines on Choosing Between a Skill or a Sub-AgentLINK
- Gap Decorations Are Now Available, Hereβs Whatβs NewLINK
- Alibabaβs AI coded for 16 days straight and every commit is on GitHubLINK
- How Stripe Built Kai on Deep Agents in 1 WeekLINK
- Valve publicly releases Lepton and FEX compatibility tools to power Steam Frame gamingLINK
- How to build a cloud software factory - computer use verificationLINK
- Smaller, faster, safer: running Kimi and GLM at scaleLINK
- Alibaba shares rally after unveiling its 'most powerful' AI model as U.S.-China competition heats upLINK
- Your agent needs a computer, not a container, introducing @cloudflare/computerLINK
- Azure Private Link for Elastic Cloud Serverless is now generally availableLINK
- Core Ubuntu package to be turned into a Snap; Deb ditchedLINK
π§° Trending tools
Prefactor: an evaluation layer that scores agent runs in real time, catching quality regressions and drift before they impact production customers.LINK
SKI: lets you voice-code with Claude Code, Codex and other agents, hearing spoken responses back so you build hands-free at thinking speed.LINK
Humalike x Hermes: gives AI agents turn-taking, timing, and memory APIs so they know when to speak, wait, or interrupt naturally.LINK
Openbase: helps developers pick reliable open-source packages by comparing popularity, activity, and reliability metrics alongside real user reviews.LINK
Pushary: sends AI agent approvals and questions to your phone's lock screen, letting Claude Code, Cursor, or Codex keep working while you're away.LINK
Claude Code usage tracking by LangWatch: monitors token usage, costs, cache efficiency, and tool calls across coding sessions with full terminal replay, free for individuals via npx.LINK
π Trending papers & reports
Cross-benchmark skill training shows that teaching an AI agent on 363 unrelated tasks still boosted its performance on five separate outside tests by roughly 3 to 10 percentage points, even on tasks like software engineering it never practiced, suggesting it learned genuinely transferable problem-solving habits rather than just memorizing test patterns.LINK
Code authorship detection nails identifying programmers from coding-contest submissions, hitting 92.6% accuracy for 10 authors, but collapses to near-zero on real classroom coursework, undermining its use for catching student cheating.LINK
Android app mapping combines code analysis with smart automated exploration to test apps, covering ~18% more of an app using 345 fewer interactions than standard random testing, speeding up quality checks before release.LINK
Android app testing maps an app's internal screen logic in detail, generating tests that cover ~16% more of the app's code while taking ~84% less time than existing tools.LINK
App testing tool maps an app's full menu structure so its automated tester skips random guessing, reaching complete screen coverage roughly 10x faster than today's random-tap testing tools.LINK
See you tomorrow for a new dose of βοΈ Devshot!