mirror of
https://github.com/snapotter-hq/SnapOtter.git
synced 2026-08-03 07:46:42 +02:00
* fix(docs): keep gray-matter on js-yaml 3 so the docs site builds The js-yaml >=4.2.0 override from #257 forced js-yaml 4 onto gray-matter (used by vitepress and vitepress-plugin-llms), which calls the removed yaml.safeLoad and broke `vitepress build`. Scope a gray-matter>js-yaml ^3.14.1 override so gray-matter keeps the v3 API (build-time, trusted frontmatter only) while app code stays on js-yaml 4.2.0+. * docs: add per-tool reference pages for all 157 tools, with a modality sidebar Generate /tools/<id> pages for the 104 tools that lacked one (video 29, audio 17, document 36, data 10, and 12 newer image tools), matching the existing page format (API endpoint, parameters from the OpenAPI spec, curl example, response, notes). Async/AI tools document the 202+SSE flow and feature-bundle requirement. Sidebar: add Video / Audio / PDF & Documents / Data groups with per-tool links, fold the 12 new image tools into the existing image categories, and replace the placeholder rest.md-anchor group. Docs site builds cleanly (157 pages, no dead links).
48 lines
1.4 KiB
Markdown
48 lines
1.4 KiB
Markdown
---
|
|
description: Split a CSV into smaller files by row count.
|
|
---
|
|
|
|
# Split CSV
|
|
|
|
Split a large CSV or TSV file into smaller files by row count. Returns a ZIP archive containing the parts.
|
|
|
|
## API Endpoint
|
|
|
|
`POST /api/v1/tools/split-csv`
|
|
|
|
Accepts multipart form data with a CSV file and a JSON `settings` field.
|
|
|
|
## Parameters
|
|
|
|
| Parameter | Type | Required | Default | Description |
|
|
|-----------|------|----------|---------|-------------|
|
|
| rowsPerFile | integer | No | `1000` | Number of data rows per output file (1--1,000,000) |
|
|
| keepHeader | boolean | No | `true` | Repeat the header row in each output file |
|
|
|
|
## Example Request
|
|
|
|
```bash
|
|
curl -X POST http://localhost:1349/api/v1/tools/split-csv \
|
|
-H "Authorization: Bearer si_your-api-key" \
|
|
-F "file=@large-dataset.csv" \
|
|
-F 'settings={"rowsPerFile": 500, "keepHeader": true}'
|
|
```
|
|
|
|
## Example Response
|
|
|
|
```json
|
|
{
|
|
"jobId": "a1b2c3d4-e5f6-7890-abcd-ef1234567890",
|
|
"downloadUrl": "/api/v1/download/a1b2c3d4-e5f6-7890-abcd-ef1234567890/large-dataset-split.zip",
|
|
"originalSize": 1048576,
|
|
"processedSize": 1050000
|
|
}
|
|
```
|
|
|
|
## Notes
|
|
|
|
- Output is always a ZIP archive containing the split CSV parts, named sequentially (e.g. `part-1.csv`, `part-2.csv`).
|
|
- When `keepHeader` is `true`, each part includes the original header row so each file can be used independently.
|
|
- Both CSV and TSV files are accepted as input.
|
|
- The row count refers to data rows only; the header row is not counted.
|