unstructured 0.27.16
Update5 days agoUnverifiedAdded Oct 10, 2026
This release rolls up all changes since 0.27.10; versions 0.27.11–0.27.15 were not published to PyPI individually. Fixes Reject images whose frames decode to too many pixels instead of exhausting memory.
Topics: Streaming
More unstructured releases
Every unstructured releaseunstructured 0.27.22
This release rolls up all changes since 0.27.16. Fixes - Faster chunk serialization with stable source-element IDs. Avoid copying orig_elements twice during serialization. IDs minted during serialization now remain stable across calls, including inside orig_elements; explicit and hash-derived IDs are unchanged.
unstructured 0.27.25
- fix(pptx): no page break before the first slide when starting_page_number > 1 - fix(chunking): reject a negative overlap instead of corrupting text - feat(image): partition WEBP images like PNG and JPEG
unstructured 0.27.10
- fix(html): extract definition lists instead of discarding them
unstructured 0.27.9
- fix(chunking): retain table rowspans across cell splits
unstructured 0.27.8
- fix: partition multi-section DOCX files in linear time - fix: ignore processing instructions in HTML partitioning - ci: remove release version alert - fix: tolerate non-UTF-8 soffice output in doc/ppt conversion - fix: recognize single-block list items as ListItem - perf(nlp): run the spaCy pipeline once per distinct text - fix(html): preserve table...
unstructured 0.27.6
- Fix DOCX text_as_html duplicating merged-cell text instead of colspan/rowspan
unstructured 0.27.5
- chore(deps): bump claude-code-action to v1 - Added Sign up link for Transform and removed pricing - fix: stop treating every hyphen as a bullet delimiter in text partitioning - fix: preserve attachment elements' own filetype in auto.partition() - chore: make a release
unstructured 0.27.1
- fix: support core metadata 2.5 publishing
Also shipped on Oct 5, 2026
Unstructured in October 2026Scheduled Search and Reporting is now generally available
New Relic is excited to share that Scheduled Search and Reporting is now generally available. Instead of manually re-running the same NRQL query to check on your data, Scheduled Search runs it for you, on whatever cadence you choose, and delivers the results straight to your inbox. What is Scheduled Search and Reporting?
Starting on July 1, 2026, Identity Service for GKE is deprecated in GKE version 1.36 and earlier
Starting on July 1, 2026, Identity Service for GKE is deprecated in GKE version 1.36 and earlier. This feature is also unavailable in organizations that were created on or after July 1, 2025. GKE version 1.37 and later don't support Identity Service for GKE.
GKE support for using the c4-standard-* machine types (up to 192 vCPUs) as Confidential GKE Nodes with Intel...
GKE support for using the c4-standard-* machine types (up to 192 vCPUs) as Confidential GKE Nodes with Intel TDX is generally available. For more information, see the following pages: To use this feature with GKE , see Encrypt workload data in-use with Confidential GKE Nodes .
AI Deep Scan with linked repositories | GitHub Cloud | Essentials and above, AI Deep Scan access, Beta
AI Deep Scan with linked repositories | You can now give AI Deep Scan context from linked repositories to analyze your primary repository more thoroughly. Source from dependencies, such as a shared authentication library, helps CodeRabbit investigate and verify security findings in the primary repository. Learn more in the linked repositories documentation .
Lyria 3 music models for your app's AI features
Your app's AI features can now generate music with two Google models, from a text prompt or an image, with vocals or as an instrumental: Lyria 3 Clip Preview , for clips of about 30 seconds. Lyria 3 Pro Preview , for full tracks up to 184 seconds. Ask Lovable to use either model in your app, or describe what you want and let Lovable pick.
iOS 27.2 beta 3 (24B5099f)
- View downloads View release notes