Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

Right. High temperature and thermal throttling are things that can and should be monitored directly.

Spotting it by digging down from service-level monitoring instead isn't terrible, but it isn't something to be proud of either.



It’s just the org chart being reflected in their stack architecture. There is a group responsible for machine health. That group doesn’t run any services. There is a group that operates GFEs. They don’t own any machines.


Does that imply a SaaS-based revision to Conway's law?

(GFEs are a reverse proxy that terminates TCP for Google. https://landing.google.com/sre/sre-book/chapters/production-... )


maybe something to be proud of by the person who actually did the digging down.




Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: