<?xml version='1.0' encoding='UTF-8'?>
<?xml-stylesheet href="/static/style.xsl" type="text/xsl"?>
<rss xmlns:atom="http://www.w3.org/2005/Atom" xmlns:content="http://purl.org/rss/1.0/modules/content/" version="2.0">
  <channel>
    <title>Most recent entries from all</title>
    <link>https://vulnerability.circl.lu</link>
    <description>Contains only the most 10 recent entries.</description>
    <docs>http://www.rssboard.org/rss-specification</docs>
    <generator>python-feedgen</generator>
    <language>en</language>
    <lastBuildDate>Thu, 01 Oct 2026 21:39:30 +0000</lastBuildDate>
    <item>
      <title>GHSA-6hm5-jgcp-p838 — Natural Language Toolkit (NLTK): Path Traversal in NKJPCorpusReader leads to Arbitrary File Read and bypasses the nltk.…</title>
      <link>https://vulnerability.circl.lu/vuln/ghsa-6hm5-jgcp-p838</link>
      <description>&lt;p&gt;&lt;strong&gt;Affected:&lt;/strong&gt; PyPI: nltk&lt;/p&gt;
&lt;p&gt;### Summary
   A path-traversal vulnerability in `NKJPCorpusReader` allows an attacker who can
   influence the `fileids` argument of its public read methods (`header`, `raw`,
   `words`, `sents`, `tagged_words`) to read files outside the corpus root. The
   reader builds the file path with no containment check and opens it with the
   builtin `open()`, so it bypasses NLTK&amp;#39;s `nltk.pathsec` sandbox — including the
   strict `ENFORCE = True` mode that `SECURITY.md` recommends for web/multi-tenant
   deployments. `header()` returns the parsed content of the out-of-root file to
   the caller (arbitrary file read).&lt;/p&gt;
&lt;p&gt;### Details
   `SECURITY.md` promises that file access is &amp;#34;validated against allowed NLTK data
   directories&amp;#34; and that with `nltk.pathsec.ENFORCE = True` &amp;#34;unauthorized file
   access … will raise `PermissionError`.&amp;#34; That guarantee is enforced via
   `FileSystemPathPointer.open()` / `CorpusReader.open()`, which call
   `nltk.pathsec.validate_path(...)`.&lt;/p&gt;
&lt;p&gt;`NKJPCorpusReader` never uses that protected path. In
   `nltk/corpus/reader/nkjp.py`:&lt;/p&gt;
&lt;p&gt;- `add_root()` builds the path by **plain string concatenation** with no
     normalization or containment check:
     ```python
     def add_root(self, fileid):          # lines 96-102
         if self.root in fileid:
             return fileid                # attacker-controlled value returned unchanged
         return self.root + fileid        # plain concat, &amp;#39;..&amp;#39; not stripped
     ```
   - The header view appends a fi…&lt;/p&gt;</description>
      <content:encoded>&lt;p&gt;&lt;strong&gt;Affected:&lt;/strong&gt; PyPI: nltk&lt;/p&gt;
&lt;p&gt;### Summary
   A path-traversal vulnerability in `NKJPCorpusReader` allows an attacker who can
   influence the `fileids` argument of its public read methods (`header`, `raw`,
   `words`, `sents`, `tagged_words`) to read files outside the corpus root. The
   reader builds the file path with no containment check and opens it with the
   builtin `open()`, so it bypasses NLTK&amp;#39;s `nltk.pathsec` sandbox — including the
   strict `ENFORCE = True` mode that `SECURITY.md` recommends for web/multi-tenant
   deployments. `header()` returns the parsed content of the out-of-root file to
   the caller (arbitrary file read).&lt;/p&gt;
&lt;p&gt;### Details
   `SECURITY.md` promises that file access is &amp;#34;validated against allowed NLTK data
   directories&amp;#34; and that with `nltk.pathsec.ENFORCE = True` &amp;#34;unauthorized file
   access … will raise `PermissionError`.&amp;#34; That guarantee is enforced via
   `FileSystemPathPointer.open()` / `CorpusReader.open()`, which call
   `nltk.pathsec.validate_path(...)`.&lt;/p&gt;
&lt;p&gt;`NKJPCorpusReader` never uses that protected path. In
   `nltk/corpus/reader/nkjp.py`:&lt;/p&gt;
&lt;p&gt;- `add_root()` builds the path by **plain string concatenation** with no
     normalization or containment check:
     ```python
     def add_root(self, fileid):          # lines 96-102
         if self.root in fileid:
             return fileid                # attacker-controlled value returned unchanged
         return self.root + fileid        # plain concat, &amp;#39;..&amp;#39; not stripped
     ```
   - The header view appends a fi…&lt;/p&gt;</content:encoded>
      <guid isPermaLink="false">https://vulnerability.circl.lu/vuln/ghsa-6hm5-jgcp-p838</guid>
    </item>
    <item>
      <title>PYSEC-2026-3581 — Natural Language Toolkit (NLTK): Path Traversal in NKJPCorpusReader leads to Arbitrary File Read and bypasses the nltk.…</title>
      <link>https://vulnerability.circl.lu/vuln/pysec-2026-3581</link>
      <description>&lt;p&gt;&lt;strong&gt;Affected:&lt;/strong&gt; PyPI: nltk&lt;/p&gt;
&lt;p&gt;### Summary
   A path-traversal vulnerability in `NKJPCorpusReader` allows an attacker who can
   influence the `fileids` argument of its public read methods (`header`, `raw`,
   `words`, `sents`, `tagged_words`) to read files outside the corpus root. The
   reader builds the file path with no containment check and opens it with the
   builtin `open()`, so it bypasses NLTK&amp;#39;s `nltk.pathsec` sandbox — including the
   strict `ENFORCE = True` mode that `SECURITY.md` recommends for web/multi-tenant
   deployments. `header()` returns the parsed content of the out-of-root file to
   the caller (arbitrary file read).&lt;/p&gt;
&lt;p&gt;### Details
   `SECURITY.md` promises that file access is &amp;#34;validated against allowed NLTK data
   directories&amp;#34; and that with `nltk.pathsec.ENFORCE = True` &amp;#34;unauthorized file
   access … will raise `PermissionError`.&amp;#34; That guarantee is enforced via
   `FileSystemPathPointer.open()` / `CorpusReader.open()`, which call
   `nltk.pathsec.validate_path(...)`.&lt;/p&gt;
&lt;p&gt;`NKJPCorpusReader` never uses that protected path. In
   `nltk/corpus/reader/nkjp.py`:&lt;/p&gt;
&lt;p&gt;- `add_root()` builds the path by **plain string concatenation** with no
     normalization or containment check:
     ```python
     def add_root(self, fileid):          # lines 96-102
         if self.root in fileid:
             return fileid                # attacker-controlled value returned unchanged
         return self.root + fileid        # plain concat, &amp;#39;..&amp;#39; not stripped
     ```
   - The header view appends a fi…&lt;/p&gt;</description>
      <content:encoded>&lt;p&gt;&lt;strong&gt;Affected:&lt;/strong&gt; PyPI: nltk&lt;/p&gt;
&lt;p&gt;### Summary
   A path-traversal vulnerability in `NKJPCorpusReader` allows an attacker who can
   influence the `fileids` argument of its public read methods (`header`, `raw`,
   `words`, `sents`, `tagged_words`) to read files outside the corpus root. The
   reader builds the file path with no containment check and opens it with the
   builtin `open()`, so it bypasses NLTK&amp;#39;s `nltk.pathsec` sandbox — including the
   strict `ENFORCE = True` mode that `SECURITY.md` recommends for web/multi-tenant
   deployments. `header()` returns the parsed content of the out-of-root file to
   the caller (arbitrary file read).&lt;/p&gt;
&lt;p&gt;### Details
   `SECURITY.md` promises that file access is &amp;#34;validated against allowed NLTK data
   directories&amp;#34; and that with `nltk.pathsec.ENFORCE = True` &amp;#34;unauthorized file
   access … will raise `PermissionError`.&amp;#34; That guarantee is enforced via
   `FileSystemPathPointer.open()` / `CorpusReader.open()`, which call
   `nltk.pathsec.validate_path(...)`.&lt;/p&gt;
&lt;p&gt;`NKJPCorpusReader` never uses that protected path. In
   `nltk/corpus/reader/nkjp.py`:&lt;/p&gt;
&lt;p&gt;- `add_root()` builds the path by **plain string concatenation** with no
     normalization or containment check:
     ```python
     def add_root(self, fileid):          # lines 96-102
         if self.root in fileid:
             return fileid                # attacker-controlled value returned unchanged
         return self.root + fileid        # plain concat, &amp;#39;..&amp;#39; not stripped
     ```
   - The header view appends a fi…&lt;/p&gt;</content:encoded>
      <guid isPermaLink="false">https://vulnerability.circl.lu/vuln/pysec-2026-3581</guid>
    </item>
  </channel>
</rss>
