{
  "title": "PDF to audio",
  "human_url": "https://happytails.ai/en/speech-tools/pdf-to-audio/",
  "agent_url": "https://happytails.ai/agents/en/speech-tools/pdf-to-audio/",
  "language": "en",
  "markdown_url": "https://happytails.ai/en/speech-tools/pdf-to-audio.md",
  "content": "PDF to audio\nRead selectable PDF text aloud using a local system voice.\nUseful next steps\nOCR PDF\nAdd searchable English text to scanned PDF pages.\nText to speech\nRead text aloud with a browser voice.\nDocument summarizer\nSummarize readable document text with source references.\nHow to use PDF to audio\nChoose a file in PDF.\nReview the settings, then choose Convert file.\nReview the result, then use Copy result or a download link when available.\nWhat to expect\nRead selectable PDF text aloud using a local system voice. Processed locally on this website’s server without an external conversion API. Uses the installed local macOS voice. Up to 6,000 extracted characters. Scans need OCR first. Linux needs a separate local voice engine.",
  "content_format": "text/plain",
  "visibility": "public",
  "links": [
    {
      "name": "OCR PDF Add searchable English text to scanned PDF pages.",
      "url": "https://happytails.ai/en/pdf-tools/ocr-pdf/"
    },
    {
      "name": "Text to speech Read text aloud with a browser voice.",
      "url": "https://happytails.ai/en/speech-tools/text-to-speech/"
    },
    {
      "name": "Document summarizer Summarize readable document text with source references.",
      "url": "https://happytails.ai/en/document-tools/document-summarizer/"
    }
  ],
  "tool": {
    "id": "pdf-to-audio",
    "category": "speech-tools",
    "built": true,
    "requirements": "Extract PDF reading order with preview; choose sections; generate chaptered audio. OCR optional for scans; tables and multi-column layouts need review before synthesis.",
    "implementation": {
      "fields": [
        {
          "key": "rate",
          "label": "Reading speed (words per minute)",
          "value": 170,
          "type": "number",
          "min": 80,
          "max": 300,
          "step": 1
        }
      ],
      "files": true,
      "accept": ".pdf",
      "multiple": false,
      "input": false,
      "processing": "Server required",
      "native": true,
      "extra": true,
      "maxFileBytes": 20971520,
      "note": "Processed locally on this website’s server without an external conversion API. Uses the installed local macOS voice. Up to 6,000 extracted characters. Scans need OCR first. Linux needs a separate local voice engine.",
      "experience": {
        "automatic": false,
        "single": false,
        "label": "Your text",
        "placeholder": "Type or paste your input…",
        "help": "Up to 1,000,000 characters. Review the settings before running.",
        "action": "Convert file"
      },
      "mode": "explicit",
      "interactive": false
    },
    "execution": "POST /api/v1/tools/pdf-to-audio/run",
    "api": {
      "id": "pdf-to-audio",
      "name": "PDF to audio",
      "category": "speech-tools",
      "description": "Read selectable PDF text aloud using a local system voice.",
      "notes": "Processed locally on this website’s server without an external conversion API. Uses the installed local macOS voice. Up to 6,000 extracted characters. Scans need OCR first. Linux needs a separate local voice engine.",
      "human_path": "/en/speech-tools/pdf-to-audio/",
      "method": "POST",
      "endpoint": "/api/v1/tools/pdf-to-audio/run",
      "schema_url": "/api/v1/tools/pdf-to-audio",
      "input_schema": {
        "type": "object",
        "additionalProperties": false,
        "properties": {
          "input": {
            "type": "string",
            "maxLength": 1000000,
            "default": "",
            "description": "Plain text input. File tools use files instead unless otherwise documented."
          },
          "options": {
            "type": "object",
            "additionalProperties": false,
            "properties": {
              "rate": {
                "type": "number",
                "description": "Reading speed (words per minute)",
                "default": 170.0,
                "minimum": 80,
                "maximum": 300,
                "multipleOf": 1
              }
            }
          },
          "files": {
            "type": "array",
            "maxItems": 1,
            "items": {
              "type": "object",
              "required": [
                "name",
                "base64"
              ],
              "additionalProperties": false,
              "properties": {
                "name": {
                  "type": "string",
                  "maxLength": 200,
                  "description": "Filename only, no path."
                },
                "mime": {
                  "type": "string",
                  "maxLength": 150
                },
                "bytes": {
                  "type": "integer",
                  "minimum": 0,
                  "description": "Optional decoded byte count; must match content if supplied."
                },
                "base64": {
                  "type": "string",
                  "contentEncoding": "base64",
                  "description": "File bytes as padded base64. Use the tool-specific limits.files_bytes value for the total decoded input size."
                }
              }
            }
          }
        }
      },
      "output_schema": {
        "type": "object",
        "required": [
          "tool",
          "result"
        ],
        "properties": {
          "tool": {
            "type": "string"
          },
          "result": {
            "type": "object",
            "required": [
              "text",
              "files"
            ],
            "properties": {
              "text": {
                "type": [
                  "string",
                  "null"
                ]
              },
              "data": {
                "description": "Parsed JSON when the textual result is JSON."
              },
              "name": {
                "type": "string"
              },
              "mime": {
                "type": "string"
              },
              "files": {
                "type": "array",
                "items": {
                  "type": "object",
                  "required": [
                    "name",
                    "mime",
                    "bytes",
                    "base64"
                  ],
                  "properties": {
                    "name": {
                      "type": "string"
                    },
                    "mime": {
                      "type": "string"
                    },
                    "bytes": {
                      "type": "integer"
                    },
                    "base64": {
                      "type": "string",
                      "contentEncoding": "base64"
                    }
                  }
                }
              }
            }
          }
        }
      },
      "processing": "Server; self-hosted native engine, no external conversion service",
      "limits": {
        "request_bytes": 31457280,
        "input_characters": 1000000,
        "files_bytes": 20971520,
        "max_files": 1,
        "timeout_seconds": 90
      }
    },
    "agent_processing": "Server; self-hosted native engine, no external conversion service",
    "browser_agent": {
      "supported": true,
      "name": "run_current_tool",
      "discovery": "WebMCP on the human page in a supporting browser",
      "verification": "See audit/API-AUDIT.md; availability is not verification"
    }
  },
  "markdown": "# PDF to audio\n\n- Human page: https://happytails.ai/en/speech-tools/pdf-to-audio/\n- Markdown: https://happytails.ai/en/speech-tools/pdf-to-audio.md\n- Structured JSON: https://happytails.ai/agents/en/speech-tools/pdf-to-audio/index.json\n- Visibility: public\n\n## Tool capabilities\n\n| Property | Value |\n| --- | --- |\n| Tool ID | `pdf-to-audio` |\n| Category | speech-tools |\n| Implementation | Registered implementation. Registration alone is not a production-quality guarantee. |\n| Browser execution | Use the visible action or interactive controls. |\n| Browser processing | Server required |\n| HTTP API | /api/v1/tools/pdf-to-audio/run |\n| Browser-agent interface | run_current_tool |\n\n### Implementation notes\n\nProcessed locally on this website’s server without an external conversion API. Uses the installed local macOS voice. Up to 6,000 extracted characters. Scans need OCR first. Linux needs a separate local voice engine.\n\n### Inputs and settings\n\nNo free-text input is required. Use the file inputs and/or settings below.\n\nAccepted browser files: `.pdf`. Choose one file.\n\nBrowser file-input limit: 20971520 bytes in total. HTTP limits can differ.\n\n| Setting | Meaning | Control type | Default | Constraints |\n| --- | --- | --- | --- | --- |\n| `rate` | Reading speed (words per minute) | number | 170 | min=80 max=300 step=1 |\n\n### HTTP contract\n\nAPI calls process supplied data on the server. The website origin is `https://happytails.ai`. Server API hosting is not connected on the public website yet. Use your local server origin to test these requests.\n\n```http\nPOST https://happytails.ai/api/v1/tools/pdf-to-audio/run\nContent-Type: application/json\n```\n\n[Detailed operation reference](https://happytails.ai/en/api/tools/pdf-to-audio.md) · [Shared API guide](https://happytails.ai/en/api.md)\n\n#### Limits\n\n| Limit | Value |\n| --- | --- |\n| request_bytes | 31457280 |\n| input_characters | 1000000 |\n| files_bytes | 20971520 |\n| max_files | 1 |\n| timeout_seconds | 90 |\n\n#### Complete input schema\n\n```json\n{\n  \"type\": \"object\",\n  \"additionalProperties\": false,\n  \"properties\": {\n    \"input\": {\n      \"type\": \"string\",\n      \"maxLength\": 1000000,\n      \"default\": \"\",\n      \"description\": \"Plain text input. File tools use files instead unless otherwise documented.\"\n    },\n    \"options\": {\n      \"type\": \"object\",\n      \"additionalProperties\": false,\n      \"properties\": {\n        \"rate\": {\n          \"type\": \"number\",\n          \"description\": \"Reading speed (words per minute)\",\n          \"default\": 170.0,\n          \"minimum\": 80,\n          \"maximum\": 300,\n          \"multipleOf\": 1\n        }\n      }\n    },\n    \"files\": {\n      \"type\": \"array\",\n      \"maxItems\": 1,\n      \"items\": {\n        \"type\": \"object\",\n        \"required\": [\n          \"name\",\n          \"base64\"\n        ],\n        \"additionalProperties\": false,\n        \"properties\": {\n          \"name\": {\n            \"type\": \"string\",\n            \"maxLength\": 200,\n            \"description\": \"Filename only, no path.\"\n          },\n          \"mime\": {\n            \"type\": \"string\",\n            \"maxLength\": 150\n          },\n          \"bytes\": {\n            \"type\": \"integer\",\n            \"minimum\": 0,\n            \"description\": \"Optional decoded byte count; must match content if supplied.\"\n          },\n          \"base64\": {\n            \"type\": \"string\",\n            \"contentEncoding\": \"base64\",\n            \"description\": \"File bytes as padded base64. Use the tool-specific limits.files_bytes value for the total decoded input size.\"\n          }\n        }\n      }\n    }\n  }\n}\n```\n\n#### Complete output schema\n\n```json\n{\n  \"type\": \"object\",\n  \"required\": [\n    \"tool\",\n    \"result\"\n  ],\n  \"properties\": {\n    \"tool\": {\n      \"type\": \"string\"\n    },\n    \"result\": {\n      \"type\": \"object\",\n      \"required\": [\n        \"text\",\n        \"files\"\n      ],\n      \"properties\": {\n        \"text\": {\n          \"type\": [\n            \"string\",\n            \"null\"\n          ]\n        },\n        \"data\": {\n          \"description\": \"Parsed JSON when the textual result is JSON.\"\n        },\n        \"name\": {\n          \"type\": \"string\"\n        },\n        \"mime\": {\n          \"type\": \"string\"\n        },\n        \"files\": {\n          \"type\": \"array\",\n          \"items\": {\n            \"type\": \"object\",\n            \"required\": [\n              \"name\",\n              \"mime\",\n              \"bytes\",\n              \"base64\"\n            ],\n            \"properties\": {\n              \"name\": {\n                \"type\": \"string\"\n              },\n              \"mime\": {\n                \"type\": \"string\"\n              },\n              \"bytes\": {\n                \"type\": \"integer\"\n              },\n              \"base64\": {\n                \"type\": \"string\",\n                \"contentEncoding\": \"base64\"\n              }\n            }\n          }\n        }\n      }\n    }\n  }\n}\n```\n\n#### Error and retry handling\n\nFailures return a non-200 status and `error.code` plus `error.message`. Correct invalid input before retrying; retry server-busy responses with bounded backoff. Returned files contain Base64 data, not persistent download URLs. Decode and inspect the output before treating conversion as successful. See the shared API guide for the complete status and timeout rules.\n\n### Browser-agent workflow\n\n1. Open the human page in a browser that supports WebMCP.\n2. Discover the interface exposed by that page; use its actual schema.\n3. Select files on the page first when the tool requires files.\n4. Call `run_current_tool` with valid input and settings.\n5. Inspect the returned result or error and the displayed output. Retrieve files from the displayed download links.\n\nCalls do not automatically copy or download results. Browser permissions still apply. Interface availability alone is not proof of successful execution.\n\n### Planning specification\n\nExtract PDF reading order with preview; choose sections; generate chaptered audio. OCR optional for scans; tables and multi-column layouts need review before synthesis.\n\nThis describes intended requirements. Use the implementation notes, interface schema and observed output to determine current support.\n\n### Complete machine-readable capability record\n\n```json\n{\n  \"id\": \"pdf-to-audio\",\n  \"category\": \"speech-tools\",\n  \"built\": true,\n  \"requirements\": \"Extract PDF reading order with preview; choose sections; generate chaptered audio. OCR optional for scans; tables and multi-column layouts need review before synthesis.\",\n  \"implementation\": {\n    \"fields\": [\n      {\n        \"key\": \"rate\",\n        \"label\": \"Reading speed (words per minute)\",\n        \"value\": 170,\n        \"type\": \"number\",\n        \"min\": 80,\n        \"max\": 300,\n        \"step\": 1\n      }\n    ],\n    \"files\": true,\n    \"accept\": \".pdf\",\n    \"multiple\": false,\n    \"input\": false,\n    \"processing\": \"Server required\",\n    \"native\": true,\n    \"extra\": true,\n    \"maxFileBytes\": 20971520,\n    \"note\": \"Processed locally on this website’s server without an external conversion API. Uses the installed local macOS voice. Up to 6,000 extracted characters. Scans need OCR first. Linux needs a separate local voice engine.\",\n    \"experience\": {\n      \"automatic\": false,\n      \"single\": false,\n      \"label\": \"Your text\",\n      \"placeholder\": \"Type or paste your input…\",\n      \"help\": \"Up to 1,000,000 characters. Review the settings before running.\",\n      \"action\": \"Convert file\"\n    },\n    \"mode\": \"explicit\",\n    \"interactive\": false\n  },\n  \"execution\": \"POST /api/v1/tools/pdf-to-audio/run\",\n  \"api\": {\n    \"id\": \"pdf-to-audio\",\n    \"name\": \"PDF to audio\",\n    \"category\": \"speech-tools\",\n    \"description\": \"Read selectable PDF text aloud using a local system voice.\",\n    \"notes\": \"Processed locally on this website’s server without an external conversion API. Uses the installed local macOS voice. Up to 6,000 extracted characters. Scans need OCR first. Linux needs a separate local voice engine.\",\n    \"human_path\": \"/en/speech-tools/pdf-to-audio/\",\n    \"method\": \"POST\",\n    \"endpoint\": \"/api/v1/tools/pdf-to-audio/run\",\n    \"schema_url\": \"/api/v1/tools/pdf-to-audio\",\n    \"input_schema\": {\n      \"type\": \"object\",\n      \"additionalProperties\": false,\n      \"properties\": {\n        \"input\": {\n          \"type\": \"string\",\n          \"maxLength\": 1000000,\n          \"default\": \"\",\n          \"description\": \"Plain text input. File tools use files instead unless otherwise documented.\"\n        },\n        \"options\": {\n          \"type\": \"object\",\n          \"additionalProperties\": false,\n          \"properties\": {\n            \"rate\": {\n              \"type\": \"number\",\n              \"description\": \"Reading speed (words per minute)\",\n              \"default\": 170.0,\n              \"minimum\": 80,\n              \"maximum\": 300,\n              \"multipleOf\": 1\n            }\n          }\n        },\n        \"files\": {\n          \"type\": \"array\",\n          \"maxItems\": 1,\n          \"items\": {\n            \"type\": \"object\",\n            \"required\": [\n              \"name\",\n              \"base64\"\n            ],\n            \"additionalProperties\": false,\n            \"properties\": {\n              \"name\": {\n                \"type\": \"string\",\n                \"maxLength\": 200,\n                \"description\": \"Filename only, no path.\"\n              },\n              \"mime\": {\n                \"type\": \"string\",\n                \"maxLength\": 150\n              },\n              \"bytes\": {\n                \"type\": \"integer\",\n                \"minimum\": 0,\n                \"description\": \"Optional decoded byte count; must match content if supplied.\"\n              },\n              \"base64\": {\n                \"type\": \"string\",\n                \"contentEncoding\": \"base64\",\n                \"description\": \"File bytes as padded base64. Use the tool-specific limits.files_bytes value for the total decoded input size.\"\n              }\n            }\n          }\n        }\n      }\n    },\n    \"output_schema\": {\n      \"type\": \"object\",\n      \"required\": [\n        \"tool\",\n        \"result\"\n      ],\n      \"properties\": {\n        \"tool\": {\n          \"type\": \"string\"\n        },\n        \"result\": {\n          \"type\": \"object\",\n          \"required\": [\n            \"text\",\n            \"files\"\n          ],\n          \"properties\": {\n            \"text\": {\n              \"type\": [\n                \"string\",\n                \"null\"\n              ]\n            },\n            \"data\": {\n              \"description\": \"Parsed JSON when the textual result is JSON.\"\n            },\n            \"name\": {\n              \"type\": \"string\"\n            },\n            \"mime\": {\n              \"type\": \"string\"\n            },\n            \"files\": {\n              \"type\": \"array\",\n              \"items\": {\n                \"type\": \"object\",\n                \"required\": [\n                  \"name\",\n                  \"mime\",\n                  \"bytes\",\n                  \"base64\"\n                ],\n                \"properties\": {\n                  \"name\": {\n                    \"type\": \"string\"\n                  },\n                  \"mime\": {\n                    \"type\": \"string\"\n                  },\n                  \"bytes\": {\n                    \"type\": \"integer\"\n                  },\n                  \"base64\": {\n                    \"type\": \"string\",\n                    \"contentEncoding\": \"base64\"\n                  }\n                }\n              }\n            }\n          }\n        }\n      }\n    },\n    \"processing\": \"Server; self-hosted native engine, no external conversion service\",\n    \"limits\": {\n      \"request_bytes\": 31457280,\n      \"input_characters\": 1000000,\n      \"files_bytes\": 20971520,\n      \"max_files\": 1,\n      \"timeout_seconds\": 90\n    }\n  },\n  \"agent_processing\": \"Server; self-hosted native engine, no external conversion service\",\n  \"browser_agent\": {\n    \"supported\": true,\n    \"name\": \"run_current_tool\",\n    \"discovery\": \"WebMCP on the human page in a supporting browser\",\n    \"verification\": \"See audit/API-AUDIT.md; availability is not verification\"\n  }\n}\n```\n\n## Page guide\n\nRead selectable PDF text aloud using a local system voice.\n\n## Useful next steps\n\n- [OCR PDF](https://happytails.ai/en/pdf-tools/ocr-pdf/): Add searchable English text to scanned PDF pages. ([Markdown](https://happytails.ai/en/pdf-tools/ocr-pdf.md))\n\n- [Text to speech](https://happytails.ai/en/speech-tools/text-to-speech/): Read text aloud with a browser voice. ([Markdown](https://happytails.ai/en/speech-tools/text-to-speech.md))\n\n- [Document summarizer](https://happytails.ai/en/document-tools/document-summarizer/): Summarize readable document text with source references. ([Markdown](https://happytails.ai/en/document-tools/document-summarizer.md))\n\n## How to use PDF to audio\n\n1. Choose a file in PDF.\n2. Review the settings, then choose Convert file.\n3. Review the result, then use Copy result or a download link when available.\n\n## What to expect\n\nRead selectable PDF text aloud using a local system voice. Processed locally on this website’s server without an external conversion API. Uses the installed local macOS voice. Up to 6,000 extracted characters. Scans need OCR first. Linux needs a separate local voice engine.\n"
}