{"_id":"@arbs.io/asset-extractor-wasm","_rev":"9-b8db6296816f0630c65c9b942bab61d8","name":"@arbs.io/asset-extractor-wasm","dist-tags":{"latest":"0.1.3"},"versions":{"0.0.1":{"name":"@arbs.io/asset-extractor-wasm","version":"0.0.1","license":"MIT","_id":"@arbs.io/asset-extractor-wasm@0.0.1","maintainers":[{"name":"arbs.io","email":"butsona@arbs.io"}],"homepage":"https://github.com/arbs-io/asset-extractor-wasm#readme","bugs":{"url":"https://github.com/arbs-io/asset-extractor-wasm/issues"},"dist":{"shasum":"43d76c720df98607b11640632a22dd06adc0d978","tarball":"https://registry.npmjs.org/@arbs.io/asset-extractor-wasm/-/asset-extractor-wasm-0.0.1.tgz","fileCount":6,"integrity":"sha512-qrxieFZGAOzDcjpl7px4u22i0TLg6zH9iEq5ldmgeXFWJY4BLscJHYFvB05q98TpzCmtRQpsN1pMgdAnIfz4gQ==","signatures":[{"sig":"MEUCICm4lrmHWO4yXrNKZxkNgkHYS1QxNzjBFH/SM6C0cBPGAiEAopyXpJ1+x5q2PDYgRkcGPoL0QZzObMlhEx0tE+dQL6A=","keyid":"SHA256:jl3bwswu80PjjokCgh0o2w5c2U4LhQAE57gj9cz1kzA"}],"unpackedSize":1098370},"main":"asset-extractor-wasm.js","types":"asset-extractor-wasm.d.ts","gitHead":"0a3346e2b52e0cf282dbe3a09dbeab4eb1199192","private":false,"sponsor":{"urk":"https://github.com/sponsors/arbs-io"},"_npmUser":{"name":"arbs.io","email":"butsona@arbs.io"},"publisher":"AndrewButson","repository":{"url":"git+https://github.com/arbs-io/asset-extractor-wasm.git","type":"git"},"_npmVersion":"9.8.0","description":"This npm package offers a straightforward method to extract text content from various binary and text file formats. The package comes with a pre-built configuration that works out-of-the-box, requiring no additional setup. It is designed for use in Browse","directories":{},"_nodeVersion":"18.16.0","collaborators":["arbs.io"],"_hasShrinkwrap":false,"_npmOperationalInternal":{"tmp":"tmp/asset-extractor-wasm_0.0.1_1692771890947_0.12434061089066573","host":"s3://npm-registry-packages"}},"0.0.2":{"name":"@arbs.io/asset-extractor-wasm","version":"0.0.2","license":"MIT","_id":"@arbs.io/asset-extractor-wasm@0.0.2","maintainers":[{"name":"arbs.io","email":"butsona@arbs.io"}],"homepage":"https://github.com/arbs-io/asset-extractor-wasm#readme","bugs":{"url":"https://github.com/arbs-io/asset-extractor-wasm/issues"},"dist":{"shasum":"8e198068484cceecb4c50f5a53263075748451fe","tarball":"https://registry.npmjs.org/@arbs.io/asset-extractor-wasm/-/asset-extractor-wasm-0.0.2.tgz","fileCount":6,"integrity":"sha512-MUbZ2nHV/TvqrJMsC5XH271M0PNFvRLNwn/danV5b9kqKKa7PZh1YV5sUJ7zv5/QBqcqrTfOeVFRetYACu6Xgw==","signatures":[{"sig":"MEUCIQCWX2qT3IawHnJjHp9m7sVfNBzm7/eqYil0MTztU7cSoAIgP5C+gSndd2ScGp8ruCYK04gcwH6LyKTEePA04oj1OrQ=","keyid":"SHA256:jl3bwswu80PjjokCgh0o2w5c2U4LhQAE57gj9cz1kzA"}],"unpackedSize":956149},"main":"asset-extractor-wasm.js","types":"asset-extractor-wasm.d.ts","gitHead":"857d8e6fa51602152a021c637aa7aed7c4de569f","private":false,"sponsor":{"urk":"https://github.com/sponsors/arbs-io"},"_npmUser":{"name":"arbs.io","email":"butsona@arbs.io"},"publisher":"AndrewButson","repository":{"url":"git+https://github.com/arbs-io/asset-extractor-wasm.git","type":"git"},"_npmVersion":"9.8.0","description":"This npm package offers a straightforward method to extract text content from various binary and text file formats. The package comes with a pre-built configuration that works out-of-the-box, requiring no additional setup. It is designed for use in Browse","directories":{},"_nodeVersion":"18.16.0","collaborators":["arbs.io"],"_hasShrinkwrap":false,"_npmOperationalInternal":{"tmp":"tmp/asset-extractor-wasm_0.0.2_1692861014534_0.9542577396143352","host":"s3://npm-registry-packages"}},"0.0.3":{"name":"@arbs.io/asset-extractor-wasm","version":"0.0.3","keywords":["pdf-parse","pdf-parser","pdf-extract","pdf-extractor","pdf-text-extract","to-text","pdf-to-text","docx-to-text","odt-to-text","html-to-text"],"license":"MIT","_id":"@arbs.io/asset-extractor-wasm@0.0.3","maintainers":[{"name":"arbs.io","email":"butsona@arbs.io"}],"homepage":"https://github.com/arbs-io/asset-extractor-wasm#readme","bugs":{"url":"https://github.com/arbs-io/asset-extractor-wasm/issues"},"dist":{"shasum":"7e5a888b64e556051189108a406a63b116b48c05","tarball":"https://registry.npmjs.org/@arbs.io/asset-extractor-wasm/-/asset-extractor-wasm-0.0.3.tgz","fileCount":6,"integrity":"sha512-B8oxpYpX3E4iyRjZfTpqksL2RXUeu6FouUjK5DKFVSTceVafoDMmYhIqTdTil925iKgfUAHI6F1d4qMOEfJrcA==","signatures":[{"sig":"MEUCIQCqghJU3AG8eu67mkClEoPMAhnBXIxpEfjZmD7vn4SDaQIgYtEpKbArTtGLw0L4q1ZiQME3hKOp6lU4+lhNwyiA8lU=","keyid":"SHA256:jl3bwswu80PjjokCgh0o2w5c2U4LhQAE57gj9cz1kzA"}],"unpackedSize":1222503},"main":"asset-extractor-wasm.js","types":"asset-extractor-wasm.d.ts","gitHead":"642aa7763bfaa0ed7788e9ca48fc744cedbab8a7","private":false,"sponsor":{"urk":"https://github.com/sponsors/arbs-io"},"_npmUser":{"name":"arbs.io","email":"butsona@arbs.io"},"publisher":"AndrewButson","repository":{"url":"git+https://github.com/arbs-io/asset-extractor-wasm.git","type":"git"},"_npmVersion":"9.8.0","description":"This npm package offers a straightforward method to extract text content from various binary and text file formats. The package comes with a pre-built configuration that works out-of-the-box, requiring no additional setup. It is designed for use in Browse","directories":{},"_nodeVersion":"18.16.0","collaborators":["arbs.io"],"_hasShrinkwrap":false,"_npmOperationalInternal":{"tmp":"tmp/asset-extractor-wasm_0.0.3_1693086573896_0.4136331314323116","host":"s3://npm-registry-packages"}},"0.0.4":{"name":"@arbs.io/asset-extractor-wasm","version":"0.0.4","keywords":["pdf-parse","pdf-parser","pdf-extract","pdf-extractor","pdf-text-extract","to-text","pdf-to-text","docx-to-text","odt-to-text","html-to-text"],"license":"MIT","_id":"@arbs.io/asset-extractor-wasm@0.0.4","maintainers":[{"name":"arbs.io","email":"butsona@arbs.io"}],"homepage":"https://github.com/arbs-io/asset-extractor-wasm#readme","bugs":{"url":"https://github.com/arbs-io/asset-extractor-wasm/issues"},"dist":{"shasum":"0d89cac8a4805a10b0671edd9ae03cb8f7798dbd","tarball":"https://registry.npmjs.org/@arbs.io/asset-extractor-wasm/-/asset-extractor-wasm-0.0.4.tgz","fileCount":6,"integrity":"sha512-2q15pZgTUoFKigijSRDO1ODfmebyzdh/1F1qz1S+PP817aWIzEuvMmb1vsmRwgFNnrhqUXl7FuP8X+qL3ZtYhA==","signatures":[{"sig":"MEUCIQDufZv0/EF32KtRhFQA8QUTnpRdGDCMNwCQ8aFJqFfLPgIgOedgEgaQ7lNmt1C5drOj6hE/7leyKWaJZWonO21eRL0=","keyid":"SHA256:jl3bwswu80PjjokCgh0o2w5c2U4LhQAE57gj9cz1kzA"}],"unpackedSize":1188909},"main":"asset-extractor-wasm.js","types":"asset-extractor-wasm.d.ts","gitHead":"2ac4a8f2a24d665b016fd3d03b7b22b1e31b0987","private":false,"sponsor":{"urk":"https://github.com/sponsors/arbs-io"},"_npmUser":{"name":"arbs.io","email":"butsona@arbs.io"},"publisher":"AndrewButson","repository":{"url":"git+https://github.com/arbs-io/asset-extractor-wasm.git","type":"git"},"_npmVersion":"9.8.0","description":"This npm package offers a straightforward method to extract text content from various binary and text file formats. The package comes with a pre-built configuration that works out-of-the-box, requiring no additional setup. It is designed for use in Browse","directories":{},"_nodeVersion":"18.16.0","collaborators":["arbs.io"],"_hasShrinkwrap":false,"_npmOperationalInternal":{"tmp":"tmp/asset-extractor-wasm_0.0.4_1693141415827_0.16406077088622828","host":"s3://npm-registry-packages"}},"0.0.5":{"name":"@arbs.io/asset-extractor-wasm","version":"0.0.5","keywords":["pdf-parse","pdf-parser","pdf-extract","pdf-extractor","pdf-text-extract","to-text","pdf-to-text","docx-to-text","odt-to-text","html-to-text"],"license":"MIT","_id":"@arbs.io/asset-extractor-wasm@0.0.5","maintainers":[{"name":"arbs.io","email":"butsona@arbs.io"}],"homepage":"https://github.com/arbs-io/asset-extractor-wasm#readme","bugs":{"url":"https://github.com/arbs-io/asset-extractor-wasm/issues"},"dist":{"shasum":"e6c776be839b949c9afdb4ad6051089edc6a10b7","tarball":"https://registry.npmjs.org/@arbs.io/asset-extractor-wasm/-/asset-extractor-wasm-0.0.5.tgz","fileCount":6,"integrity":"sha512-GPdRg60DNgmCasGAJzLOKQAQ+F+bn+MXSyoiI41nMNO1HZiaGpBCz/vqnTqkHbhszMK621NSyt/etff/GWFPuQ==","signatures":[{"sig":"MEUCIQDs9FlI2+LyA+IBC11lIptcQzfIvwffSnDgawGcv6u0fgIgdFK5/SCWDux+ngftUwVJWOFN5cGM9UqZeNiqIfBQYwA=","keyid":"SHA256:jl3bwswu80PjjokCgh0o2w5c2U4LhQAE57gj9cz1kzA"}],"unpackedSize":1703207},"main":"asset-extractor-wasm.js","types":"asset-extractor-wasm.d.ts","gitHead":"5b854c47098daf30e8c37b1b827b3b02a0449958","private":false,"sponsor":{"urk":"https://github.com/sponsors/arbs-io"},"_npmUser":{"name":"arbs.io","email":"butsona@arbs.io"},"publisher":"AndrewButson","repository":{"url":"git+https://github.com/arbs-io/asset-extractor-wasm.git","type":"git"},"_npmVersion":"9.8.0","description":"This npm package offers a straightforward method to extract text content from various binary and text file formats. The package comes with a pre-built configuration that works out-of-the-box, requiring no additional setup. It is designed for use in Browse","directories":{},"_nodeVersion":"18.16.0","collaborators":["arbs.io"],"_hasShrinkwrap":false,"_npmOperationalInternal":{"tmp":"tmp/asset-extractor-wasm_0.0.5_1693251109892_0.012335173045471493","host":"s3://npm-registry-packages"}},"0.0.6":{"name":"@arbs.io/asset-extractor-wasm","version":"0.0.6","keywords":["pdf-parse","pdf-parser","pdf-extract","pdf-extractor","pdf-text-extract","to-text","pdf-to-text","docx-to-text","odt-to-text","html-to-text"],"license":"MIT","_id":"@arbs.io/asset-extractor-wasm@0.0.6","maintainers":[{"name":"arbs.io","email":"butsona@arbs.io"}],"homepage":"https://github.com/arbs-io/asset-extractor-wasm#readme","bugs":{"url":"https://github.com/arbs-io/asset-extractor-wasm/issues"},"dist":{"shasum":"914fff0205e3a4899b285ee4b81ee154650e1da2","tarball":"https://registry.npmjs.org/@arbs.io/asset-extractor-wasm/-/asset-extractor-wasm-0.0.6.tgz","fileCount":6,"integrity":"sha512-3vlcTiXGfpQBnXSs5yFvOQT44XjUoynoy2vdYKQEV76HPGRExXQfvefJhe618MCo2ZCW4LWVKTU2zyxOOPsZqg==","signatures":[{"sig":"MEUCIDSccWdAlLlBCkYusyHaWy6pGlZCynWWRfeH3r17TZ+HAiEAw/FCSzZDr7wv8ZjeeExYde99WAiDi+army7l6XgOnxw=","keyid":"SHA256:jl3bwswu80PjjokCgh0o2w5c2U4LhQAE57gj9cz1kzA"}],"unpackedSize":1865547},"main":"asset-extractor-wasm.js","types":"asset-extractor-wasm.d.ts","gitHead":"9d6ed2502f8a345f7dbe4c32a2460d850b689e74","private":false,"sponsor":{"urk":"https://github.com/sponsors/arbs-io"},"_npmUser":{"name":"arbs.io","email":"butsona@arbs.io"},"publisher":"AndrewButson","repository":{"url":"git+https://github.com/arbs-io/asset-extractor-wasm.git","type":"git"},"_npmVersion":"9.8.0","description":"This npm package offers a straightforward method to extract text content from various binary and text file formats. The package comes with a pre-built configuration that works out-of-the-box, requiring no additional setup. It is designed for use in Browse","directories":{},"_nodeVersion":"18.16.0","collaborators":["arbs.io"],"_hasShrinkwrap":false,"_npmOperationalInternal":{"tmp":"tmp/asset-extractor-wasm_0.0.6_1693774178169_0.5387920348832993","host":"s3://npm-registry-packages"}},"0.1.0":{"name":"@arbs.io/asset-extractor-wasm","version":"0.1.0","keywords":["pdf-parse","pdf-parser","pdf-extract","pdf-extractor","pdf-text-extract","to-text","pdf-to-text","docx-to-text","odt-to-text","html-to-text"],"license":"MIT","_id":"@arbs.io/asset-extractor-wasm@0.1.0","maintainers":[{"name":"arbs.io","email":"butsona@arbs.io"}],"homepage":"https://github.com/arbs-io/asset-extractor-wasm#readme","bugs":{"url":"https://github.com/arbs-io/asset-extractor-wasm/issues"},"dist":{"shasum":"0dd9b7a7e918bd4e24e15863f93c923e7dd04875","tarball":"https://registry.npmjs.org/@arbs.io/asset-extractor-wasm/-/asset-extractor-wasm-0.1.0.tgz","fileCount":6,"integrity":"sha512-7IVAPCUkXJrsEczQa0KYQy0CnPYNiwDUXSoQGjTmZkmqi7tQM+uULui3skEbCMdGQKB/M4z/S0jTqxI/pvANtw==","signatures":[{"sig":"MEUCIDlUQvc5CPbN+0sBs7ys1U+bAYsfjegLuBXeA+nJlvmsAiEAlON2+V0uqB5d2y7cCgoefY3JTbulOTVck2Gfjv+5nzI=","keyid":"SHA256:jl3bwswu80PjjokCgh0o2w5c2U4LhQAE57gj9cz1kzA"}],"unpackedSize":1465981},"main":"asset-extractor-wasm.js","types":"asset-extractor-wasm.d.ts","gitHead":"889da30bd5243f89f8b7d3ac46e3930016aa6453","private":false,"sponsor":{"urk":"https://github.com/sponsors/arbs-io"},"_npmUser":{"name":"arbs.io","email":"butsona@arbs.io"},"publisher":"AndrewButson","repository":{"url":"git+https://github.com/arbs-io/asset-extractor-wasm.git","type":"git"},"_npmVersion":"9.8.0","description":"This npm package offers a straightforward method to extract text content from various binary and text file formats. The package comes with a pre-built configuration that works out-of-the-box, requiring no additional setup. It is designed for use in Browse","directories":{},"_nodeVersion":"18.16.0","collaborators":["arbs.io"],"_hasShrinkwrap":false,"_npmOperationalInternal":{"tmp":"tmp/asset-extractor-wasm_0.1.0_1694156691139_0.5100667629145055","host":"s3://npm-registry-packages"}},"0.1.1":{"name":"@arbs.io/asset-extractor-wasm","version":"0.1.1","keywords":["pdf-parse","pdf-parser","pdf-extract","pdf-extractor","pdf-text-extract","to-text","pdf-to-text","docx-to-text","odt-to-text","html-to-text"],"license":"MIT","_id":"@arbs.io/asset-extractor-wasm@0.1.1","maintainers":[{"name":"arbs.io","email":"butsona@arbs.io"}],"homepage":"https://github.com/arbs-io/asset-extractor-wasm#readme","bugs":{"url":"https://github.com/arbs-io/asset-extractor-wasm/issues"},"dist":{"shasum":"56e9a8240786802058a5594fc334b2cd257b2086","tarball":"https://registry.npmjs.org/@arbs.io/asset-extractor-wasm/-/asset-extractor-wasm-0.1.1.tgz","fileCount":6,"integrity":"sha512-O9ktHPL37nSc87WQgfCF12+4cJB5sTKFX2QmT3WcwTj03OHes4YsU74S61QJxsu04oxVUHnRo68/2Upn0WP32g==","signatures":[{"sig":"MEYCIQD/KscQCY0vh+CdVwo4DWYLRnrG8Kij3TJQp1IpagFkywIhANKrGpfgh1xWa7VtCvMsUG6lqZtj8Y00ZRAnU8K1E7Rl","keyid":"SHA256:jl3bwswu80PjjokCgh0o2w5c2U4LhQAE57gj9cz1kzA"}],"unpackedSize":1471640},"main":"asset-extractor-wasm.js","types":"asset-extractor-wasm.d.ts","gitHead":"7c39528aa26d080064487f876a7f8bdaf18058aa","private":false,"sponsor":{"urk":"https://github.com/sponsors/arbs-io"},"_npmUser":{"name":"arbs.io","email":"butsona@arbs.io"},"publisher":"AndrewButson","repository":{"url":"git+https://github.com/arbs-io/asset-extractor-wasm.git","type":"git"},"_npmVersion":"9.8.0","description":"This npm package offers a straightforward method to extract text content from various binary and text file formats. The package comes with a pre-built configuration that works out-of-the-box, requiring no additional setup. It is designed for use in Browse","directories":{},"_nodeVersion":"18.16.0","collaborators":["arbs.io"],"_hasShrinkwrap":false,"_npmOperationalInternal":{"tmp":"tmp/asset-extractor-wasm_0.1.1_1699392909165_0.267143280609655","host":"s3://npm-registry-packages"}},"0.1.2":{"name":"@arbs.io/asset-extractor-wasm","version":"0.1.2","keywords":["pdf-parse","pdf-parser","pdf-extract","pdf-extractor","pdf-text-extract","to-text","pdf-to-text","docx-to-text","odt-to-text","html-to-text"],"license":"MIT","_id":"@arbs.io/asset-extractor-wasm@0.1.2","maintainers":[{"name":"arbs.io","email":"butsona@arbs.io"}],"homepage":"https://github.com/arbs-io/asset-extractor-wasm#readme","bugs":{"url":"https://github.com/arbs-io/asset-extractor-wasm/issues"},"dist":{"shasum":"2eba9e888f05554ba0757394a6efe5039db72e52","tarball":"https://registry.npmjs.org/@arbs.io/asset-extractor-wasm/-/asset-extractor-wasm-0.1.2.tgz","fileCount":6,"integrity":"sha512-vZI0d6NtByzFB2+hV1DfsV0Z/mKyfStZt3lsobwWw706AzepP9QYUY9Wza8/orOZ5gz8DBgVYDEUI9Ny7vbe4w==","signatures":[{"sig":"MEQCIGOyTb6DtGxaHOZbYk+WPFcL0IBXmxDqpSOhtWm1QxLbAiAgbsel2yk/unAES/Pa/eT+iKA0hsH0gKtzaA2Byu7rhA==","keyid":"SHA256:jl3bwswu80PjjokCgh0o2w5c2U4LhQAE57gj9cz1kzA"}],"unpackedSize":1402607},"main":"asset-extractor-wasm.js","types":"asset-extractor-wasm.d.ts","gitHead":"11121f808f091a6043c09cea43198a85ec1dc11c","private":false,"sponsor":{"urk":"https://github.com/sponsors/arbs-io"},"_npmUser":{"name":"arbs.io","email":"butsona@arbs.io"},"publisher":"AndrewButson","repository":{"url":"git+https://github.com/arbs-io/asset-extractor-wasm.git","type":"git"},"_npmVersion":"9.8.0","description":"This npm package offers a straightforward method to extract text content from various binary and text file formats. The package comes with a pre-built configuration that works out-of-the-box, requiring no additional setup. It is designed for use in Browse","directories":{},"_nodeVersion":"18.16.0","collaborators":["arbs.io"],"_hasShrinkwrap":false,"_npmOperationalInternal":{"tmp":"tmp/asset-extractor-wasm_0.1.2_1702545159514_0.772078920548698","host":"s3://npm-registry-packages"}},"0.1.3":{"name":"@arbs.io/asset-extractor-wasm","collaborators":["arbs.io"],"description":"This npm package offers a straightforward method to extract text content from various binary and text file formats. The package comes with a pre-built configuration that works out-of-the-box, requiring no additional setup. It is designed for use in Browse","version":"0.1.3","license":"MIT","repository":{"type":"git","url":"git+https://github.com/arbs-io/asset-extractor-wasm.git"},"main":"asset-extractor-wasm.js","types":"asset-extractor-wasm.d.ts","keywords":["pdf-parse","pdf-parser","pdf-extract","pdf-extractor","pdf-text-extract","to-text","pdf-to-text","docx-to-text","odt-to-text","html-to-text"],"private":false,"publisher":"AndrewButson","sponsor":{"urk":"https://github.com/sponsors/arbs-io"},"_id":"@arbs.io/asset-extractor-wasm@0.1.3","gitHead":"4600d07664b089e613462dc07b57046c8a3dc9b1","bugs":{"url":"https://github.com/arbs-io/asset-extractor-wasm/issues"},"homepage":"https://github.com/arbs-io/asset-extractor-wasm#readme","_nodeVersion":"23.5.0","_npmVersion":"11.0.0","dist":{"integrity":"sha512-oq+hl7+1+32sAIVyHeHA2UcloqkwsW66W+cyl86nu4ykR4JPwjjO9Mjv4cUYAt7OALCKwwWjVjEmEF3YRaGiVQ==","shasum":"b1157883dec1217d77b902476a2450f508b74dab","tarball":"https://registry.npmjs.org/@arbs.io/asset-extractor-wasm/-/asset-extractor-wasm-0.1.3.tgz","fileCount":6,"unpackedSize":1582993,"signatures":[{"keyid":"SHA256:jl3bwswu80PjjokCgh0o2w5c2U4LhQAE57gj9cz1kzA","sig":"MEUCIQDAehb2mxQnwRbkJCclGpfNVgMC6syXqUGpSn8wWNaghgIgc0S1AqGWX0WTSE4EqM3pzK4U7gdXHrcCrxcmgrAK4Rs="}]},"_npmUser":{"name":"arbs.io","email":"butsona@arbs.io"},"directories":{},"maintainers":[{"name":"arbs.io","email":"butsona@arbs.io"}],"_npmOperationalInternal":{"host":"s3://npm-registry-packages-npm-production","tmp":"tmp/asset-extractor-wasm_0.1.3_1737122754796_0.46887769790410494"},"_hasShrinkwrap":false}},"time":{"created":"2023-08-23T06:24:50.856Z","modified":"2025-01-17T14:05:55.252Z","0.0.1":"2023-08-23T06:24:51.198Z","0.0.2":"2023-08-24T07:10:15.007Z","0.0.3":"2023-08-26T21:49:34.225Z","0.0.4":"2023-08-27T13:03:36.118Z","0.0.5":"2023-08-28T19:31:50.201Z","0.0.6":"2023-09-03T20:49:38.438Z","0.1.0":"2023-09-08T07:04:51.398Z","0.1.1":"2023-11-07T21:35:09.629Z","0.1.2":"2023-12-14T09:12:39.759Z","0.1.3":"2025-01-17T14:05:55.067Z"},"bugs":{"url":"https://github.com/arbs-io/asset-extractor-wasm/issues"},"license":"MIT","homepage":"https://github.com/arbs-io/asset-extractor-wasm#readme","keywords":["pdf-parse","pdf-parser","pdf-extract","pdf-extractor","pdf-text-extract","to-text","pdf-to-text","docx-to-text","odt-to-text","html-to-text"],"repository":{"type":"git","url":"git+https://github.com/arbs-io/asset-extractor-wasm.git"},"description":"This npm package offers a straightforward method to extract text content from various binary and text file formats. The package comes with a pre-built configuration that works out-of-the-box, requiring no additional setup. It is designed for use in Browse","maintainers":[{"name":"arbs.io","email":"butsona@arbs.io"}],"readme":"# asset-extractor-wasm\n\n> **Caution**: This package is currently in development and should be treated as a **preview** release (pre-v1.0)\n\nWelcome to `@arbs.io/asset-extractor-wasm`, a powerful npm package that provides a straightforward method to extract content from a wide range of binary and text file formats. This package is pre-configured to work seamlessly, requiring no additional setup. It is designed to be compatible with both Browsers and Node.js environments, including Visual Studio Code extensions, making it a versatile tool for your development needs.\n\n## Features\n\n### Supported File Types\n\nThe current version of the package supports content extraction from an extensive list of MIME types, including but not limited to:\n\n| Text | Media | extension | Mimetype                                                                  |\n| ---- | ----- | --------- | ------------------------------------------------------------------------- |\n| ✅   | ⚫    | **txt**   | text/plain                                                                |\n| ✅   | ✅    | **docx**  | application/vnd.openxmlformats-officedocument.wordprocessingml.document   |\n| ✅   | ✅    | **pptx**  | application/vnd.openxmlformats-officedocument.presentationml.presentation |\n| 🔲   | 🔲    | **xlsx**  | application/vnd.openxmlformats-officedocument.spreadsheetml.sheet         |\n| ✅   | ✅    | **odp**   | application/vnd.oasis.opendocument.presentation                           |\n| ✅   | ✅    | **ods**   | application/vnd.oasis.opendocument.spreadsheet                            |\n| ✅   | ✅    | **odt**   | application/vnd.oasis.opendocument.text                                   |\n| ✅   | 🔲    | **xml**   | text/xml                                                                  |\n| ✅   | 🔲    | **pdf**   | application/pdf                                                           |\n| ✅   | 🔲    | **html**  | text/html                                                                 |\n| ✅   | 🔲    | **epub**  | application/epub+zip                                                      |\n| ✅   | 🔲    | **mobi**  | application/x-mobipocket-ebook                                            |\n\n#### Legend\n\n✅: Completed\n🔲: Coming soon\n⚫: Not Applicable\n\n- Text: Extract Text\n- Media: Extract Image/Video/Audio\n\n## Requesting Additional File Support\n\nWe are always looking to expand the capabilities of `@arbs.io/asset-extractor-wasm`. If you need support for additional file formats, please submit an enhancement issue on the project's repository. We value your feedback and contributions as they help us improve this package for the broader developer community.\n\n## Installation\n\nTo install the package, use the following npm command:\n\n```sh\nnpm install @arbs.io/asset-extractor-wasm\n```\n\nThis command will add the package to your project's dependencies.\n\n## Usage\n\nHere's an example of how to extract text from a buffer. If the file type is binary, the mime-type is verified using file-type.\n\n```ts\nimport * as fs from 'fs'\nimport {\n  createDocumentParser,\n  getTextPlain,\n} from '@arbs.io/asset-extractor-wasm'\n\nexport const documentParserExample = () => {\n  const buf = fs.readFileSync(`./data_source/microservices.docx`)\n  const documentParser = createDocumentParser(new Uint8Array(buf))\n\n  console.log(`mimetype: (${documentParser?.mimetype})`)\n  console.log(`extension: (${documentParser?.extension})`)\n  console.log(`content [text/plain]: (${documentParser?.contents?.text!})`)\n}\n```\n\nThis example demonstrates how to read a file, convert it to a `Uint8Array`, and then extract the assets.\n\n## API\n\n### DocumentParser\n\nThe `DocumentParser` object provides the following properties:\n\n- `mimetype`: The mime-type of the buffer determined by the binary signature.\n- `extension`: The (file) extension of the buffer determined by the binary signature.\n- `contents`: An array of `Content` within the buffer (text, images, ...)\n\n```ts\ninterface DocumentParser {\n  mimetype: string\n  extension: string\n  contents: ParserContent | null\n}\n```\n\n### ParserContent\n\n- `text`: Text content of the buffer. There is only ever a single text content for each buffer.\n- `media`: Array of all embedded media assets with the buffer (images, audio, video, ...).\n\n```ts\ninterface ParserContent {\n  text: string | null\n  media: ContentData[] | null\n}\n```\n\n### ContentData\n\n- `identity`: The identity of the binary embedded object. For example: `image1.png`\n- `mimetype`: The mime-type is set to the format of the data send to the function. For example: `image/png`\n- `data`: The raw data base64 of the image binary format\n\n```ts\ninterface ContentData {\n  identity: string\n  mimetype: string\n  data: string\n}\n```\n\nWe hope you find `@arbs.io/asset-extractor-wasm` useful for your projects. If you have any questions, issues, or suggestions, please feel free to open an issue on our GitHub repository. We appreciate your support and are committed to making this package even better for the developer community.\n","readmeFilename":"README.md"}