video-crawler: the six gates between a dropped URL and a published item
Our takeDownloading at 1080p and publishing at 720p looks contradictory; it is really quality and bandwidth billed separately, since a platform's own 720p rendition is already a second-generation encode. Covers come from sampling frames at 10%, 33%, 60% and 80% and taking the first non-black one. Two incident-born rules are worth copying: read back everything you wrote (a quoted SQL update once failed silently and the Chinese site served English titles), and launch batches of four or more with setsid fully detached, because SSH drops at five to ten minutes and nohup is unreliable.
robotworld-ingest v2.3.0 is the front door for external content: the user drops a Twitter/X, YouTube or arXiv URL and says collect it, and this skill turns it into something publishable. Of its six stages three are run by scripts (media capture, the two webp derivatives, online verification) and three must be run by an agent (metadata enrichment, paper detail pages, social assets). Video is downloaded at 1080p and published at 720p.