Agent skill

Toxy

by nissl-lab in nissl-lab/toxy

A skill your agent uses whenever the user wants to extract text or data from documents using Toxy (the .NET text extraction library).

Apache-2.0Auto-check passedDocuments & Office

Install Toxy

skills CLI
$ npx skills add nissl-lab/toxy --skill toxy -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install nissl-lab/toxy toxy --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/nissl-lab/toxy.git skills-src && mkdir -p .claude/skills && cp -r skills-src/.claude/skills/toxy .claude/skills/toxy && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
toxy
GitHub stars
463
Token cost
~1.7k tokens
SKILL.md length
356 words
Files
2 (incl. references)
Skills in repo
1
Repo updated
First seen
Licence
Apache-2.0

At a glance

A skill your agent uses whenever the user wants to extract text or data from documents using Toxy (the .NET text extraction library).

  • The user wants to extract text
  • SKILL.md covers Key Changes in 2.6, Installation, Core Concepts and Toxy Object Types, plus 3 more sections
  • Calls dotnet
  • Data from documents using Toxy (the .NET text extraction library)

What it does

Toxy is an agent skill from nissl-lab/toxy. Use this skill whenever the user wants to extract text or data from documents using Toxy (the .NET text extraction library). Trigger when the user mentions reading, parsing, or extracting content from files like docx, xlsx, xls, pdf, csv, txt, epub, html, eml, vcf using C or .NET. Also trigger when the user asks about Toxy NuGet package, ToxyDocument, ToxySpreadsheet, ToxyEmail, stream parsing, or any Toxy API usage. Use this skill for any Toxy 2.6 code generation, migration from older Toxy versions, or…

Its SKILL.md is about 1.7k tokens, which your agent loads only when the skill is triggered. The skill folder holds 2 other files, including reference files (for example `references/api.md`).

It sits in Documents & Office, covering Excel spreadsheets, Word documents and CSV and tabular files. It works with Microsoft Excel, Microsoft Word, .NET and C#. The repository describes itself as: .net text extraction & export framework. The licence is Apache-2.0.

When your agent uses it

  • The user wants to extract text
  • Data from documents using Toxy (the .NET text extraction library)
  • The user mentions reading
  • Extracting content from files like docx

Example prompts

  • “/toxy”

What it can do on your machine

Read from SKILL.md and the folder at commit 99ca266. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Shell commands in SKILL.md call:

    • dotnet

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Toxy loads about 1.7k tokens when it runs, and up to ~3.1k if it reads all its reference files. Until then it costs about 136 tokens; SKILL.md has 356 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~136
When it runs · the whole SKILL.md, loaded when a task matches
~1.7k
With references · SKILL.md plus every file in references/, read only if the agent opens them
~3.1k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from nissl-lab/toxy at commit 99ca266, republished under its Apache-2.0 licence (© nissl-lab). 356 words, ~1,662 tokens.

Download SKILL.mdSave it as .claude/skills/toxy/SKILL.md (or your agent's skills folder). This skill also uses 1 other file; get the full folder from GitHub.
name
toxy
description
Use this skill whenever the user wants to extract text or data from documents using Toxy (the .NET text extraction library). Trigger when the user mentions reading, parsing, or extracting content from files like docx, xlsx, xls, pdf, csv, txt, epub, html, eml, vcf using C# or .NET. Also trigger when the user asks about Toxy NuGet package, ToxyDocument, ToxySpreadsheet, ToxyEmail, stream parsing, or any Toxy API usage. Use this skill for any Toxy 2.6 code generation, migration from older Toxy versions, or troubleshooting Toxy parsers.

Toxy 2.6 Skill

Toxy is a .NET data/text extraction framework (similar to Apache Tika for Java). It supports cross-platform text and data extraction from many popular file formats. Always use Toxy 2.6.0 (NuGet package Toxy, targeting netstandard2.0 or netstandard2.1).

Key Changes in 2.6

  • Upgraded to .NET Standard 2.1 support (in addition to 2.0)
  • Added stream-based parsing (ParserContext now accepts Stream directly)
  • Added EPUB parser (implements IDocumentParser)
  • Removed unused StreamReader references (cleaner API)
  • All parsers now live in Toxy namespace

Installation

xml
<!-- .csproj -->
<PackageReference Include="Toxy" Version="2.6.0" />

Or via CLI:

bash
dotnet add package Toxy --version 2.6.0

Core Concepts

ParserContext

The entry point for all parsing. Accepts either a file path or a Stream (new in 2.6):

csharp
// From file path
var context = new ParserContext("path/to/file.docx");

// From stream (new in 2.6)
using var stream = File.OpenRead("path/to/file.docx");
var context = new ParserContext(stream, "docx"); // must supply format hint
Parser Factory

Use ParserFactory to auto-detect format and return the correct parser:

csharp
var parser = ParserFactory.CreateDocument(context);    // for document types
var parser = ParserFactory.CreateSpreadsheet(context); // for spreadsheet types
var parser = ParserFactory.CreateEmail(context);       // for email/contact types

Toxy Object Types

ObjectDescriptionFormats
ToxyDocumentParagraphs + metadatadocx, pdf, txt, epub, html, rtf, odt
ToxySpreadsheetRows/cells per sheetxlsx, xls, csv, ods
ToxyEmailEmail fieldseml, msg
ToxyBusinessCardContact fieldsvcf
ToxyDomDOM treehtml, xml
ToxyMetadataKey/value metadataany file

Usage Patterns

Extract Text from a Word Document
csharp
using Toxy;

var context = new ParserContext("report.docx");
var parser = ParserFactory.CreateDocument(context);
ToxyDocument doc = parser.Parse();

foreach (var paragraph in doc.Paragraphs)
{
    Console.WriteLine(paragraph.Text);
}
Extract Data from Excel
csharp
using Toxy;

var context = new ParserContext("data.xlsx");
var parser = ParserFactory.CreateSpreadsheet(context);
ToxySpreadsheet sheet = parser.Parse();

foreach (var table in sheet.Tables)
{
    Console.WriteLine($"Sheet: {table.Name}");
    foreach (var row in table.Rows)
    {
        foreach (var cell in row.Cells)
        {
            Console.Write($"{cell.Value}\t");
        }
        Console.WriteLine();
    }
}
Parse from a Stream (New in 2.6)
csharp
using Toxy;

// Works with any stream source (MemoryStream, HttpResponseStream, etc.)
using var stream = File.OpenRead("document.pdf");
var context = new ParserContext(stream, "pdf");
var parser = ParserFactory.CreateDocument(context);
ToxyDocument doc = parser.Parse();
Console.WriteLine(doc.Paragraphs[0].Text);
Parse PDF
csharp
using Toxy;

var context = new ParserContext("file.pdf");
var parser = ParserFactory.CreateDocument(context);
ToxyDocument doc = parser.Parse();

foreach (var para in doc.Paragraphs)
    Console.WriteLine(para.Text);
Parse EPUB (New in 2.6)
csharp
using Toxy;

var context = new ParserContext("book.epub");
var parser = ParserFactory.CreateDocument(context);
ToxyDocument doc = parser.Parse();

foreach (var para in doc.Paragraphs)
    Console.WriteLine(para.Text);
Parse Email
csharp
using Toxy;

var context = new ParserContext("message.eml");
var parser = ParserFactory.CreateEmail(context);
ToxyEmail email = parser.Parse();

Console.WriteLine($"From: {email.From}");
Console.WriteLine($"Subject: {email.Subject}");
Console.WriteLine($"Body: {email.Body}");
Parse Business Card (VCF)
csharp
using Toxy;

var context = new ParserContext("contact.vcf");
var parser = ParserFactory.CreateEmail(context); // VCF uses email parser factory
ToxyBusinessCard card = (ToxyBusinessCard)parser.Parse();

Console.WriteLine(card.FullName);
Console.WriteLine(card.Email);
Extract Metadata
csharp
using Toxy;

var context = new ParserContext("file.pdf");
var parser = ParserFactory.CreateMetadata(context);
ToxyMetadata meta = parser.Parse();

foreach (var key in meta.Keys)
    Console.WriteLine($"{key}: {meta[key]}");
Parse HTML as DOM
csharp
using Toxy;

var context = new ParserContext("page.html");
var parser = ParserFactory.CreateDom(context);
ToxyDom dom = parser.Parse();

// Access DOM nodes
Console.WriteLine(dom.Root.InnerText);

Show full SKILL.md (160 more words)Show less

Supported Formats Summary

FormatExtension(s)Parser Type
Word (Open XML).docxDocument
Word (Legacy).docDocument
PDF.pdfDocument
Plain Text.txtDocument
Rich Text.rtfDocument
EPUB.epubDocument (new in 2.6)
HTML.html, .htmDocument / Dom
OpenDocument Text.odtDocument
Excel (Open XML).xlsxSpreadsheet
Excel (Legacy).xlsSpreadsheet
CSV.csvSpreadsheet
OpenDocument Sheet.odsSpreadsheet
Email.eml, .msgEmail
Business Card.vcfEmail (returns ToxyBusinessCard)
Any*Metadata

Tips & Best Practices

  • Auto-detection: When using a file path, Toxy detects format from the extension automatically. When using a stream, always provide the format hint string (e.g., "pdf", "docx").
  • Error handling: Wrap parse calls in try/catch — unsupported formats throw NotSupportedException.
  • Large files: Use stream-based parsing to avoid loading entire files into memory.
  • Cross-platform: Toxy targets netstandard2.0/2.1, so it works on Windows, Linux, and macOS.
  • No IFilter dependency: Unlike old Windows-based approaches, Toxy does not require IFilter COM components.

For deeper reference on specific parsers and the class hierarchy, see references/api.md.

© nissl-lab, Apache-2.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 1 other file (references) in .claude/skills/toxy of nissl-lab/toxy.

  • SKILL.md
  • references/api.md

Open the folder on GitHubat commit 99ca266

Compare with similar skills

Toxy next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Toxy compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Toxy this skillnissl-lab/toxy463—~1.7kAutomated safety check: PassApache-2.0
Documentszhongkaifu/TensorSharp559—~4.2kAutomated safety check: PassBSD-3-Clause
Markdown Exporterbowenliang123/markdown-exporter2721 repos~5.3kAutomated safety check: PassApache-2.0
Office ArtifactsPrismer-AI/PrismerCloud1.6k—~2.6kAutomated safety check: PassMIT
MinerU Document Readeropendatalab/MinerU81k—~9.4kAutomated safety check: WarnCustom licence
File ReadingWide-Moat/open-computer-use1261 repos~3.1kAutomated safety check: PassProprietary

Similar skills

  • Documents

    zhongkaifu/TensorSharp

    Read and write real documents on the device - PDF, XLSX, DOCX, PPTX and CSV.

    559 GitHub stars~4.2k tokensUpdated today
    Documents & OfficeAuto-check passed
  • Markdown Exporter

    bowenliang123/markdown-exporter

    Convert Markdown text to DOCX, PPTX, XLSX, PDF, PNG, SVG, HTML, IPYNB, MD, CSV, JSON, JSONL, XML files, and extract code blocks in Markdown to Python, Bash,JS and etc files.

    272 GitHub starsUsed in 1 repo~5.3k tokens
    Documents & OfficeAuto-check passed
  • Office Artifacts

    Prismer-AI/PrismerCloud

    Generate real DOCX, PPTX, XLSX, PDF, CSV files using python-docx / python-pptx / openpyxl / reportlab by writing them into the dispatch artifacts dir, then explicitly deliver each one with cloud…

    1.6k GitHub stars~2.6k tokensUpdated 9 days ago
    Documents & OfficeAuto-check passed
  • MinerU Document Reader

    opendatalab/MinerU

    Reads, OCRs, searches and cites local documents through the mineru CLI, covering PDF, images, Office files, EPUB, HTML and CSV.

    81k GitHub stars~9.4k tokensUpdated today
    Documents & OfficeAuto-check: warnings
  • File Reading

    Wide-Moat/open-computer-use

    A skill your agent uses when a file has been uploaded but its content is NOT in your context — only its path at /mnt/user-data/uploads/ is listed in an uploadedfiles block.

    126 GitHub starsUsed in 1 repo~3.1k tokens
    Documents & OfficeAuto-check passed
  • Light File Reading

    Light0305/Light-skills

    Light 多格式文件深度理解常驻技能:强大地读 Word / PDF / PPTX / Excel / CSV / 图片 / 视频 / 代码 / 压缩包,不只提取文字,而是理解结构 / 图表 / 数据 / 格式要求 / 隐含意图,产结构化"理解笔记"五面 (结构逻辑·关键内容·格式约束·视觉风格·可复用)并映射到下游技能动作(这个文件→接下来能做什么)。

    640 GitHub stars~4.1k tokensUpdated 3 mo ago
    Documents & OfficeAuto-check passed

Questions about Toxy

What does Toxy do?

A skill your agent uses whenever the user wants to extract text or data from documents using Toxy (the .NET text extraction library). Toxy is an agent skill from nissl-lab/toxy.NET text extraction library).

When should I use Toxy?

Toxy fits situations like: the user wants to extract text; data from documents using Toxy (the .NET text extraction library); the user mentions reading; extracting content from files like docx.

How do I install Toxy in Claude Code?

Run `npx skills add nissl-lab/toxy --skill toxy -a claude-code`. Or copy the skill folder (.claude/skills/toxy in nissl-lab/toxy) into .claude/skills/toxy in your project. Claude Code loads it when a task matches its description.

How do I install Toxy in Codex?

Run `npx skills add nissl-lab/toxy --skill toxy -a codex`. Or copy the skill folder (.claude/skills/toxy in nissl-lab/toxy) into .agents/skills/toxy in your project. Codex loads it when a task matches its description.

Can I use Toxy in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add nissl-lab/toxy --skill toxy -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/toxy, .gemini/skills/toxy, .github/skills/toxy and .opencode/skills/toxy in your project.

What does Toxy need to run?

Going by SKILL.md and its folder, Toxy needs the command-line tools its instructions call (dotnet).

Does Toxy access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Toxy safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Toxy use?

Toxy is published under the Apache-2.0 licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Toxy use?

About 1.7k tokens (SKILL.md is roughly 6.6k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 1.4k tokens, read only when the agent opens those files.

What are the alternatives to Toxy?

Skills that share tags, products or a category with Toxy: Documents (zhongkaifu/TensorSharp, 559 stars), Markdown Exporter (bowenliang123/markdown-exporter, 272 stars), Office Artifacts (Prismer-AI/PrismerCloud, 1.6k stars) and MinerU Document Reader (opendatalab/MinerU, 81k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Toxy?

nissl-lab (a GitHub organization) maintains it in nissl-lab/toxy, which has 463 GitHub stars. The repository was last updated on June 19, 2026.

Source: nissl-lab/toxy on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.