> For clean Markdown of any page, append .md to the page URL. > For a complete documentation index, see https://docs.scrapybara.com/sdk-reference/python/llms.txt. > For AI client integration (Claude Code, Cursor, etc.), connect to the MCP server at https://docs.scrapybara.com/_mcp/server. # Python SDK [![pypi](https://img.shields.io/pypi/v/scrapybara)](https://pypi.python.org/pypi/scrapybara) #### [GitHub repository](https://github.com/scrapybara/scrapybara-python) View the Python SDK source code #### [PyPI package](https://pypi.org/project/scrapybara/) View the Scrapybara package on PyPI The Scrapybara Python library provides convenient access to the Scrapybara API from Python. ## Installation ```sh pip install scrapybara ``` ## Reference Please only refer to this documentation site for reference. The GitHub reference [here](https://github.com/scrapybara/scrapybara-python/blob/master/reference.md) is generated programmatically and incomplete. ## Usage Instantiate and use the client with the following: ```python from scrapybara import Scrapybara client = Scrapybara( api_key="YOUR_API_KEY", ) instance = client.start_ubuntu() ``` ## Async Client The SDK also exports an `async` client so that you can make non-blocking calls to our API. ```python import asyncio from scrapybara import AsyncScrapybara client = AsyncScrapybara( api_key="YOUR_API_KEY", ) async def main() -> None: await client.start_ubuntu() asyncio.run(main()) ``` ## Exception Handling When the API returns a non-success status code (4xx or 5xx response), a subclass of the following error will be thrown. ```python from scrapybara.core.api_error import ApiError try: client.start_ubuntu(...) except ApiError as e: print(e.status_code) print(e.body) ``` ## Advanced ### Retries The SDK is instrumented with automatic retries with exponential backoff. A request will be retried as long as the request is deemed retriable and the number of retry attempts has not grown larger than the configured retry limit (default: 2). A request is deemed retriable when any of the following HTTP status codes is returned: * [408](https://developer.mozilla.org/en-US/docs/Web/HTTP/Status/408) (Timeout) * [429](https://developer.mozilla.org/en-US/docs/Web/HTTP/Status/429) (Too Many Requests) * [5XX](https://developer.mozilla.org/en-US/docs/Web/HTTP/Status/500) (Internal Server Errors) Use the `max_retries` request option to configure this behavior. ```python client.start_ubuntu(..., request_options={ "max_retries": 1 }) ``` ### Timeouts The SDK defaults to a 60 second timeout. You can configure this with a timeout option at the client or request level. ```python from scrapybara import Scrapybara client = Scrapybara( ..., timeout=20.0, ) # Override timeout for a specific method client.start_ubuntu(..., request_options={ "timeout_in_seconds": 1 }) ``` ### Custom Client You can override the `httpx` client to customize it for your use-case. Some common use-cases include support for proxies and transports. ```python import httpx from scrapybara import Scrapybara client = Scrapybara( ..., httpx_client=httpx.Client( proxies="http://my.test.proxy.example.com", transport=httpx.HTTPTransport(local_address="0.0.0.0"), ), ) ``` ## Contributing While we value open-source contributions to this SDK, this library is generated programmatically. Additions made directly to this library would have to be moved over to our generation code, otherwise they would be overwritten upon the next generated release. Feel free to open a PR as a proof of concept, but know that we will not be able to merge it as-is. We suggest opening an issue first to discuss with us! On the other hand, contributions to the README are always very welcome!