Jina Reader returned 404/garbage for many news sites. Now fetch the page HTML directly, parse og:title/og:description + author + article<time> + <p> paragraphs; Jina only as fallback. Verified: theverge AI page → real title 'Artificial Intelligence', 8 paragraphs.