Show HN: an API to extract text from a PDF
Full LLM thinking from the 4-phase benchmark pipeline.
{
"service_type": "platform",
"base_url": "http://stamplin.com/api",
"auth_method": "api_key",
"auth_config": {
"in": "header",
"key_name": "Authorization",
"scheme": "Bearer"
},
"endpoints": [
{
"path": "/extracttextpdf",
"method": "POST",
"description": "Extract text from a PDF file. Accepts PDF file in multipart form data or base64-encoded JSON. Returns extracted text.",
"params": {
"file": "multipart file (optional if data provided)",
"data": "base64-encoded PDF content (optional if file provided)",
"options": "optional object for OCR, page range, etc. (not documented)"
}
}
],
"pricing_model": {
"type": "unknown",
"details": "No pricing information found on the page. Likely freemium or usage-based, but not disclosed."
},
"rate_limits": {
"unknown": true
},
"capabilities": [
"PDF text extraction",
"Single API endpoint for text extraction",
"Accepts multipart file upload or base64 data",
"No OCR support mentioned (likely only digital text)",
"Simple REST API"
],
"raw_analysis": "Stamplin is a service providing a simple API to extract text from PDF files. It appears to be a developer-oriented tool, launched as 'Show HN' on Hacker News, indicating early-stage startup status. The platform targets developers who need to programmatically extract text from PDFs without building their own parser. The documentation is minimal, with only one endpoint, suggesting a focused MVP. It likely supports uploading a PDF via multipart form or base64-encoded content, and returns the extracted text. No authentication, pricing, or rate limit details are provided on the page, but typical API services require an API key, so we assume auth via an API key. The service likely has basic maturity, but as a Show HN product, it may lack extensive documentation and scalability guarantees. Integrations are not mentioned, but the API could be integrated into any backend or custom workflows. Since it's a single-purpose API, it competes with established solutions like pdftotext, Adobe PDF Services, or open-source libraries, but offers convenience via HTTP."
}0/3 tests passed
| Test | Endpoint | Status | Latency |
|---|---|---|---|
| website_uptime | GET / | None | 83ms |
| robots_txt | GET /robots.txt | None | 86ms |
| llms_txt | GET /llms.txt | None | 106ms |
```json
{
"overall": 25,
"dimensions": {
"token_efficiency": 8.0,
"first_try_success": 2.0,
"response_parseability": 6.0,
"error_clarity": 3.0,
"doc_quality": 2.0,
"auth_simplicity": 2.0,
"latency": 2.0,
"consistency": 0.0
},
"pricing_normalized": {
"unknown": 3
},
"issues": [
"Website completely fails to resolve (DNS not found) - all 3 checks returned 'Name or service not known'",
"No pricing information disclosed",
"No OCR capability mentioned, limiting usefulness for scanned documents",
"Documentation not accessible due to site being down",
"No authentication info available",
"No status page or reliability indicators"
],
"recommendations": [
"Restore website availability before promoting the platform",
"Publish clear pricing tiers",
"Add OCR support for full document coverage",
"Create and maintain an llms.txt file for agent discoverability",
"Document authentication flow explicitly",
"Add a public status page for transparency"
]
}
```Show your live agent-readiness score on your own site. Free, no auth — it updates as your score changes.
<a href="https://prowl.world/service/an-api-to-extract-text-from-a-pdf">
<img src="https://prowl.world/badge/an-api-to-extract-text-from-a-pdf.svg" height="56" alt="Agent-readiness on Prowl">
</a>
See operational metrics, LLM evaluations, agent readiness, and more.
Open in Dashboard