Linux workstation

Upstream software project

beautifulsoup4

HTML/XML Parser for Quick-Turnaround Applications Like Screen-Scraping

About beautifulsoup4

HTML/XML Parser for Quick-Turnaround Applications Like Screen-Scraping

This project links 11 native package records across 5 recorded operating-system releases. Compare the retained versions and architectures below, then open the package for your own release.

These are catalog observations, not a guarantee of installation, compatibility, or upstream support.

Project pictures and package coverage

Debian 12 (Bookworm): 2 package records; Debian 13 (Trixie): 2 package records; openSUSE Leap 15.6: 2 package records; openSUSE Leap 16.0: 2 package records; openSUSE Tumbleweed: 3 package records. Catalog coverage diagram, not an application screenshot.beautifulsoup4: recorded package coverageDebian 12 (Bookworm)2 recordsDebian 13 (Trixie)2 recordsopenSUSE Leap 15.62 recordsopenSUSE Leap 16.02 recordsopenSUSE Tumbleweed3 records
OpenFactory diagram of linked package records. It is not an application screenshot.

Project identity

Project
beautifulsoup4
Publisher
Not authoritatively mapped
Native package records
11
Operating systems
debian-12, debian-13, opensuse-leap-15-6, opensuse-leap-16-0, opensuse-tumbleweed
License expression
MIT
Metadata completeness
100/100 (not a software quality rating)
Source repository
Not reported

Source-reported description

The fullest retained description is shown with its source. Distribution packaging descriptions may include downstream details.

Beautiful Soup is a Python HTML/XML parser designed for quick turnaround projects like screen-scraping. Three features make it powerful: * Beautiful Soup won't choke if you give it bad markup. It yields a parse tree that makes approximately as much sense as your original document. This is usually good enough to collect the data you need and run away * Beautiful Soup provides a few simple methods and Pythonic idioms for navigating, searching, and modifying a parse tree: a toolkit for dissecting a document and extracting what you need. You don't have to create a custom parser for each application * Beautiful Soup automatically converts incoming documents to Unicode and outgoing documents to UTF-8. You don't have to think about encodings, unless the document doesn't specify an encoding and Beautiful Soup can't autodetect one. Then you just have to specify the original encoding Beautiful Soup parses anything you give it, and does the tree traversal stuff for you. You can tell it "Find all the links", or "Find all the links of class externalLink", or "Find all the links whose urls match "foo.com", or "Find the table heading that's got bold text, then give me that text." Valuable data that was once locked up in poorly-designed websites is now within your reach. Projects that would have taken hours take only minutes with Beautiful Soup.

Description source

Packages by operating system

Compare recorded versions, then open a package for dependency, file, checksum, and repository evidence. Version strings are distribution-specific, not a ranking of newer software.

Debian 12 (Bookworm)

  1. python3-bs4

    Debian 12 (Bookworm) / python / source beautifulsoup4

    4.11.2-2

    error-tolerant HTML parser for Python 3

    allbookworm
  2. python-bs4-doc

    Debian 12 (Bookworm) / doc / source beautifulsoup4

    4.11.2-2

    error-tolerant HTML parser for Python - documentation

    allbookworm

Debian 13 (Trixie)

  1. python3-bs4

    Debian 13 (Trixie) / python / source beautifulsoup4

    4.13.4-2

    error-tolerant HTML parser for Python 3

    alltrixie
  2. python-bs4-doc

    Debian 13 (Trixie) / doc / source beautifulsoup4

    4.13.4-2

    error-tolerant HTML parser for Python - documentation

    alltrixie

openSUSE Leap 15.6

  1. python311-beautifulsoup4

    openSUSE Leap 15.6 / Unspecified / source python-beautifulsoup4

    4.12.2-150400.7.3.9

    HTML/XML Parser for Quick-Turnaround Applications Like Screen-Scraping

    noarchleap-15.6
  2. python3-beautifulsoup4

    openSUSE Leap 15.6 / Unspecified / source python-beautifulsoup4

    4.8.2-1.18

    HTML/XML Parser for Quick-Turnaround Applications Like Screen-Scraping

    noarchleap-15.6

openSUSE Leap 16.0

  1. python313-beautifulsoup4

    openSUSE Leap 16.0 / Unspecified / source python-beautifulsoup4

    4.12.3-160000.2.2

    HTML/XML Parser for Quick-Turnaround Applications Like Screen-Scraping

    noarchleap-16.0
  2. python-beautifulsoup4-doc

    openSUSE Leap 16.0 / Unspecified / source python-beautifulsoup4

    4.12.3-160000.2.2

    Documentation for python-beautifulsoup4

    noarchleap-16.0

openSUSE Tumbleweed

  1. python313-beautifulsoup4

    openSUSE Tumbleweed / Unspecified / source python-beautifulsoup4

    4.15.0-1.3

    HTML/XML Parser for Quick-Turnaround Applications Like Screen-Scraping

    noarchtumbleweed
  2. python314-beautifulsoup4

    openSUSE Tumbleweed / Unspecified / source python-beautifulsoup4

    4.15.0-1.3

    HTML/XML Parser for Quick-Turnaround Applications Like Screen-Scraping

    noarchtumbleweed
  3. python-beautifulsoup4-doc

    openSUSE Tumbleweed / Unspecified / source python-beautifulsoup4

    4.15.0-1.3

    Documentation for python-beautifulsoup4

    noarchtumbleweed

Project resources and further reading

Mapping provenance

Only source-backed identity signals create public cross-OS links. A reviewer can later approve or dispute an inferred relationship without rewriting native package history.

No field-level source record is published yet.