> This page is for version v1 (default).
> For other versions, use one of these documentation indexes:
> - v1 (default): https://docs.plextera.com/v-1/llms.txt

> For clean Markdown of any page, append .md to the page URL.
> For a complete documentation index, see https://docs.plextera.com/llms.txt.
> For AI client integration (Claude Code, Cursor, etc.), connect to the MCP server at https://docs.plextera.com/_mcp/server.

# List extractions

GET https://api.plextera.com/api/public/v1/document-insights/extractions

Returns a page of extraction summaries. Use **Get extraction** for full `output`.

Filter with `labels[key]=value`; multiple labels use AND.
See [Filter by labels](/guides/core-guides/document-extraction#filter-by-labels) for examples.


Reference: https://docs.plextera.com/api/api-reference/document-insights/list-document-extractions

## Authentication

- `Authorization` header (required) — Use the API key format: `api-key <token>`.

## Request

### Query parameters

- `page` (integer, optional, default: 0) — Zero-based page index. Defaults to 0.
- `size` (integer, optional, default: 50) — Page size between 1 and 100. Defaults to 50.
- `status` (enum, optional) — Filter by extraction status.
  - Allowed values: `QUEUED`, `PROCESSING`, `COMPLETED`, `FAILED`, `REJECTED`
- `from` (datetime, optional) — Return extractions created at or after this UTC ISO-8601 timestamp.
- `to` (datetime, optional) — Return extractions created at or before this UTC ISO-8601 timestamp.
- `sortBy` (enum, optional) — Field used to sort the list. Defaults to createdAt.
  - Allowed values: `createdAt`, `updatedAt`, `completedAt`
- `sort` (enum, optional) — Sort direction. Defaults to desc.
  - Allowed values: `asc`, `desc`
- `labels` (map from string to string, optional) — Filter with `labels[key]=value`. - An empty value matches extractions that contain the label. - Multiple labels are combined with AND. - Example: `?labels[docType]=YOUR_CONFIGURED_DOCUMENT_TYPE&labels[customerDocumentId]=doc-42`.

## Response

### 200

One page of extractions, without the full `output`.

- `data` (list of DocumentExtraction, optional) — Items on the requested page.
- `pageInfo` (PageInfo, optional) — Paging details for this result set.

## Errors

### 400 Bad Request Error

INVALID_REQUEST — The request is malformed: unreadable JSON, an invalid parameter value, or a bad multipart body.

- `error` (ApiError, optional) — Machine-readable error details shared by every error response.

### 401 Unauthorized Error

UNAUTHORIZED — The API key is missing, malformed, or revoked.

- `error` (ApiError, optional) — Machine-readable error details shared by every error response.

### 403 Forbidden Error

FORBIDDEN — The API key is valid but not allowed to perform this operation.

- `error` (ApiError, optional) — Machine-readable error details shared by every error response.

### 405 Method Not Allowed Error

METHOD_NOT_ALLOWED — The HTTP method is not supported for this path.

- `error` (ApiError, optional) — Machine-readable error details shared by every error response.

### 406 Not Acceptable Error

NOT_ACCEPTABLE — The requested response media type is not supported.

- `error` (ApiError, optional) — Machine-readable error details shared by every error response.

### 409 Conflict Error

CONFLICT — The request conflicts with existing state in an upstream Plextera service.

- `error` (ApiError, optional) — Machine-readable error details shared by every error response.

### 429 Too Many Requests Error

RATE_LIMITED — Too many requests in a short window. 429 responses are enforced at the network edge; the body may be plain text instead of this envelope, so rely on the status code and the Retry-After header.

- `error` (ApiError, optional) — Machine-readable error details shared by every error response.

### 500 Internal Server Error

INTERNAL_ERROR — Unexpected failure inside Plextera. Retry with backoff; contact support with the requestId if it persists.

- `error` (ApiError, optional) — Machine-readable error details shared by every error response.

### 502 Bad Gateway Error

DOWNSTREAM_ERROR — A downstream Plextera service failed while handling the request. Retry with backoff.

- `error` (ApiError, optional) — Machine-readable error details shared by every error response.

## Types

### DocumentExtraction

Document Insights extraction summary, without the full `output`.

- `id` (string, optional) — Extraction identifier. Store this value for polling, feedback, and event correlation. Example: `69654f0bc073ef404baec649`.
- `status` (enum, optional) — Current processing state. `FAILED` covers routing, processing, or extracted-data validation problems. `REJECTED` covers document-level problems such as duplicate content or an unsupported language. Stop polling when the value is `COMPLETED`, `FAILED`, or `REJECTED`.
  - Allowed values: `QUEUED`, `PROCESSING`, `COMPLETED`, `FAILED`, `REJECTED`
- `outputAvailable` (boolean, optional) — `true` means the final `output` is ready and included in the extraction details or completed event payload.
- `createdAt` (datetime, optional) — UTC timestamp when Plextera accepted the extraction request. Example: `2026-04-07T10:05:04Z`.
- `updatedAt` (datetime, optional) — UTC timestamp of the latest known state change. Example: `2026-04-07T10:22:00Z`.
- `completedAt` (datetime, optional) — UTC timestamp when the extraction reached `COMPLETED`, `FAILED`, or `REJECTED`. Omitted while the extraction is still queued or processing.
- `error` (RunError, optional) — Failure or rejection details. Present only when `status` is `FAILED` or `REJECTED`.
- `labels` (map from string to string, optional) — String key-value pairs submitted with the extraction and returned for correlation.
- `document` (DocumentInsightsDocument, optional) — Summary of the source document that was processed.

### PageInfo

Paging details returned by every list endpoint.

- `page` (integer, optional) — Zero-based index of the returned page.
- `size` (integer, optional) — Requested page size.
- `totalItems` (long, optional) — Total number of items matching the query.
- `totalPages` (integer, optional) — Total number of pages for the requested size.

### ApiError

Machine-readable error details shared by every error response.

- `code` (string, optional) — Stable machine-readable error code. Branch your handling on this, not on `message`.
- `message` (string, optional) — Human-readable explanation. Wording may change; do not parse it.
- `requestId` (string, optional) — Unique request identifier, also returned in the X-Request-Id response header. Include it in support requests.
- `retryable` (boolean, optional) — True when retrying the same request later may succeed.
- `details` (list of ApiErrorDetail, optional) — Field-level validation issues. Present only for VALIDATION_FAILED.

### RunError

Terminal failure details. Present only for failed or rejected extractions and workflow runs.

- `code` (string, optional) — Stable machine-readable reason code. Document Insights returns `DOCUMENT_NOT_ROUTED` with `FAILED` for routing problems, `DOCUMENT_EXTRACTION_FAILED` with `FAILED` for processing or extracted-data validation problems, and `DOCUMENT_REJECTED` with `REJECTED` for document-level problems such as duplicate content or an unsupported language.
- `message` (string, optional) — Human-readable explanation of the failure or rejection.

### DocumentInsightsDocument

Document summary used inside Document Insights extraction and event payloads. Includes file metadata returned by File Service and page count returned by Document Insights; optional fields are omitted if the upstream service does not return them.

- `fileId` (string, optional) — Plextera File Service identifier for the source document.
- `fileName` (string, optional) — Source file name returned by File Service. Omitted when the upstream file record does not include a name.
- `mimeType` (string, optional) — Standard MIME media type of the source document, for example `application/pdf`. Omitted when Document Insights cannot determine the media type.
- `size` (long, optional) — Source document size in bytes returned by File Service. Omitted when file size is unavailable.
- `pageCount` (integer, optional) — Source document page count returned by Document Insights. Omitted when page count has not been computed.
- `contentUrl` (string, optional, nullable) — Short-lived pre-signed URL to download the source document directly from File Service. Returned only by Get extraction (`GET /extractions/{id}`); it is minted per request and expires, so use it promptly and do not store it. Omitted from List extractions responses and whenever the download URL cannot be resolved.

### ApiErrorDetail

Field-level validation issue.

- `field` (string, optional) — Name of the invalid request field.
- `issue` (string, optional) — What is wrong with the field value.

## Examples

**Response**

```json
{
  "data": [
    {
      "id": "69654f0bc073ef404baec649",
      "status": "COMPLETED",
      "outputAvailable": true,
      "createdAt": "2026-04-07T10:05:04Z",
      "updatedAt": "2026-04-07T10:22:00Z",
      "completedAt": "2026-04-07T10:22:00Z",
      "labels": {
        "customerDocumentId": "doc-42",
        "docType": "invoice"
      },
      "document": {
        "fileId": "file_01JY7M4ZVX5R1P3M3Q0TA1S7ZM",
        "fileName": "invoice.pdf",
        "mimeType": "application/pdf",
        "size": 63877,
        "pageCount": 1
      }
    }
  ],
  "pageInfo": {
    "page": 0,
    "size": 50,
    "totalItems": 1,
    "totalPages": 1
  }
}
```

**SDK Code**

```python Completed extractions page
import requests

url = "https://api.plextera.com/api/public/v1/document-insights/extractions"

querystring = {"from":"2026-04-01T00:00:00Z","labels":"{\"docType\":\"YOUR_CONFIGURED_DOCUMENT_TYPE\"}","page":"0","size":"50","sort":"desc","sortBy":"createdAt","status":"COMPLETED","to":"2026-04-30T23:59:59Z"}

headers = {"Authorization": "<apiKey>"}

response = requests.get(url, headers=headers, params=querystring)

print(response.json())
```

```javascript Completed extractions page
const url = 'https://api.plextera.com/api/public/v1/document-insights/extractions?from=2026-04-01T00%3A00%3A00Z&labels=%7B%22docType%22%3A%22YOUR_CONFIGURED_DOCUMENT_TYPE%22%7D&page=0&size=50&sort=desc&sortBy=createdAt&status=COMPLETED&to=2026-04-30T23%3A59%3A59Z';
const options = {method: 'GET', headers: {Authorization: '<apiKey>'}};

try {
  const response = await fetch(url, options);
  const data = await response.json();
  console.log(data);
} catch (error) {
  console.error(error);
}
```

```go Completed extractions page
package main

import (
	"fmt"
	"net/http"
	"io"
)

func main() {

	url := "https://api.plextera.com/api/public/v1/document-insights/extractions?from=2026-04-01T00%3A00%3A00Z&labels=%7B%22docType%22%3A%22YOUR_CONFIGURED_DOCUMENT_TYPE%22%7D&page=0&size=50&sort=desc&sortBy=createdAt&status=COMPLETED&to=2026-04-30T23%3A59%3A59Z"

	req, _ := http.NewRequest("GET", url, nil)

	req.Header.Add("Authorization", "<apiKey>")

	res, _ := http.DefaultClient.Do(req)

	defer res.Body.Close()
	body, _ := io.ReadAll(res.Body)

	fmt.Println(res)
	fmt.Println(string(body))

}
```

```ruby Completed extractions page
require 'uri'
require 'net/http'

url = URI("https://api.plextera.com/api/public/v1/document-insights/extractions?from=2026-04-01T00%3A00%3A00Z&labels=%7B%22docType%22%3A%22YOUR_CONFIGURED_DOCUMENT_TYPE%22%7D&page=0&size=50&sort=desc&sortBy=createdAt&status=COMPLETED&to=2026-04-30T23%3A59%3A59Z")

http = Net::HTTP.new(url.host, url.port)
http.use_ssl = true

request = Net::HTTP::Get.new(url)
request["Authorization"] = '<apiKey>'

response = http.request(request)
puts response.read_body
```

```java Completed extractions page
import com.mashape.unirest.http.HttpResponse;
import com.mashape.unirest.http.Unirest;

HttpResponse<String> response = Unirest.get("https://api.plextera.com/api/public/v1/document-insights/extractions?from=2026-04-01T00%3A00%3A00Z&labels=%7B%22docType%22%3A%22YOUR_CONFIGURED_DOCUMENT_TYPE%22%7D&page=0&size=50&sort=desc&sortBy=createdAt&status=COMPLETED&to=2026-04-30T23%3A59%3A59Z")
  .header("Authorization", "<apiKey>")
  .asString();
```

```php Completed extractions page
<?php
require_once('vendor/autoload.php');

$client = new \GuzzleHttp\Client();

$response = $client->request('GET', 'https://api.plextera.com/api/public/v1/document-insights/extractions?from=2026-04-01T00%3A00%3A00Z&labels=%7B%22docType%22%3A%22YOUR_CONFIGURED_DOCUMENT_TYPE%22%7D&page=0&size=50&sort=desc&sortBy=createdAt&status=COMPLETED&to=2026-04-30T23%3A59%3A59Z', [
  'headers' => [
    'Authorization' => '<apiKey>',
  ],
]);

echo $response->getBody();
```

```csharp Completed extractions page
using RestSharp;

var client = new RestClient("https://api.plextera.com/api/public/v1/document-insights/extractions?from=2026-04-01T00%3A00%3A00Z&labels=%7B%22docType%22%3A%22YOUR_CONFIGURED_DOCUMENT_TYPE%22%7D&page=0&size=50&sort=desc&sortBy=createdAt&status=COMPLETED&to=2026-04-30T23%3A59%3A59Z");
var request = new RestRequest(Method.GET);
request.AddHeader("Authorization", "<apiKey>");
IRestResponse response = client.Execute(request);
```

```swift Completed extractions page
import Foundation

let headers = ["Authorization": "<apiKey>"]

let request = NSMutableURLRequest(url: NSURL(string: "https://api.plextera.com/api/public/v1/document-insights/extractions?from=2026-04-01T00%3A00%3A00Z&labels=%7B%22docType%22%3A%22YOUR_CONFIGURED_DOCUMENT_TYPE%22%7D&page=0&size=50&sort=desc&sortBy=createdAt&status=COMPLETED&to=2026-04-30T23%3A59%3A59Z")! as URL,
                                        cachePolicy: .useProtocolCachePolicy,
                                    timeoutInterval: 10.0)
request.httpMethod = "GET"
request.allHTTPHeaderFields = headers

let session = URLSession.shared
let dataTask = session.dataTask(with: request as URLRequest, completionHandler: { (data, response, error) -> Void in
  if (error != nil) {
    print(error as Any)
  } else {
    let httpResponse = response as? HTTPURLResponse
    print(httpResponse)
  }
})

dataTask.resume()
```