Πρόσθετα προγράμματος περιήγησης Firefox
  • Επεκτάσεις
  • Θέματα
    • για το Firefox
    • Λεξικά και πακέτα γλωσσών
    • Άλλα προγράμματα περιήγησης
    • Πρόσθετα για Android
Σύνδεση
Προεπισκόπηση του Batch Text Extractor (URL2Text)

Batch Text Extractor (URL2Text) από Trevor Lewis

Extract readable text from up to 100 URLs sequentially and download validated results as one or more plain-text files.

Διαθέσιμο στο Firefox για Android™Διαθέσιμο στο Firefox για Android™
5 (1 κριτική)5 (1 κριτική)
73 χρήστες73 χρήστες
Λήψη Firefox και απόκτηση επέκτασης
Λήψη αρχείου
Σαρώστε τον κωδικό QR για να ανοίξετε αυτήν την επέκταση στο Firefox για Android

Μεταδεδομένα επέκτασης

Στιγμιότυπα
Enter up to 100 URLs, choose your loading time and output options, then extract clean, source-labelled text locally in Firefox.
Σχετικά με την επέκταση
Batch Text Extractor converts lists of webpages into clean, source-labelled plain text. It is designed for researchers, analysts, journalists, content teams, and anyone who needs to collect readable webpage content efficiently.

How it works:

• Enter up to 100 URLs, one per line.
• Choose how long each page should remain open, from 0 to 10 seconds. The recommended default is 5 seconds.
• Start the batch and allow the extension to process each URL sequentially.
• Each URL opens in a temporary background tab, readable content is extracted, and the tab closes before the next URL opens.
• Results are compiled and downloaded as one to ten plain-text files.

Reliable text extraction:

• Waits for meaningful and stable content instead of relying only on a fixed timer.
• Extracts visible text from the main page, accessible frames, and open shadow roots.
• Records the requested URL and the final URL after redirects.
• Detects corrupted, encoded, repeated, or machine-generated content.
• Rejects CAPTCHA screens, login gates, access-denied pages, rate-limit responses, JavaScript shells, and server-error pages.
• Automatically retries a failed URL once using a longer loading period.
• Provides clear failure reasons when content cannot be extracted safely.
• Validates the completed output before downloading it.

Batch management:

• Detects valid, invalid, unsupported, and duplicate URLs before processing.
• Processes duplicate URLs once by default, with an option to retain and process them separately.
• Allows an optional batch name for descriptive filenames.
• Divides large batches evenly across up to ten output files while preserving URL order.
• Prevents multiple batches from running simultaneously.
• Continues processing if the extension popup is closed.
• Restores live progress when the popup is reopened.
• Provides a completion summary with successful, failed, retried, and partial extraction counts.
• Allows failed URLs to be retried without re-entering the complete batch.
• Offers optional Firefox completion notifications.

Public and authenticated pages:

The extension can process public webpages and pages where you are already signed in through your current Firefox session. It does not request, collect, or store website login credentials.

Privacy:

• All extraction happens locally within Firefox.
• No analytics or tracking.
• No advertising services.
• No external APIs or developer servers.
• No user data is collected or transmitted.
• The extension runs only when you manually start a batch.

Batch Text Extractor is intentionally focused on reliable URL-to-text extraction. It does not crawl links, generate AI summaries, perform OCR, or alter the webpages it processes.

Requires Firefox 140 or later on desktop and Firefox 142 or later on Android.
Σχόλια προγραμματιστών
Tips & Known Limitations

The extension processes URLs sequentially, opening only one temporary background tab at a time and closing it before continuing.

The selected page wait time can be set from 0 to 10 seconds. This is treated as a minimum wait. The content-readiness guard may wait up to 12 additional seconds when meaningful and stable text has not yet appeared. Five seconds is the recommended default.

If extraction fails, the extension automatically closes the temporary tab, waits briefly, and retries the URL once with a longer loading period. Automatic retries never repeat more than once.

The extension supports public webpages and pages where the user is already authenticated through the current Firefox session.

Known limitations:

• Infinite scrolling and content requiring user interaction may not be fully captured.
• Closed shadow roots, canvas text, images, video content, and OCR are not supported.
• Firefox-protected pages, internal browser pages, extension pages, PDF viewer pages, and restricted Mozilla domains cannot be scripted.
• CAPTCHA screens, login gates, access-denied pages, rate-limit responses, JavaScript-required shells, and server-error pages are recorded as failures rather than valid content.
• If accessible-frame extraction fails, the extension may use the top-level document and mark the extraction as partial.
• A maximum of 100 URLs can be processed in one batch.

Extracted content is checked for corruption, encoded data, machine-generated noise, repeated-content loops, insufficient readable language, and unstable page content. Invalid content is never presented as a successful extraction.

Bug Reports & Feedback

Spotted a bug or want to suggest a feature? Email: trevorza@outlook.com
Βαθμολογήθηκε με 5 από έναν αξιολογητή
Συνδεθείτε για βαθμολόγηση της επέκτασης
Δεν υπάρχουν ακόμη βαθμολογίες

Η βαθμολογία αστεριών αποθηκεύτηκε

5
1
4
0
3
0
2
0
1
0
Ανάγνωση 1 κριτικής
Δικαιώματα και δεδομένα

Απαιτούμενα δικαιώματα:

  • Κάνει λήψη αρχείων και ανάγνωση/τροποποίηση ιστορικού λήψεων του προγράμματος περιήγησης
  • Έχει πρόσβαση στις καρτέλες περιήγησης
  • Έχει πρόσβαση στα δεδομένα σας για κάθε ιστότοπο

Προαιρετικά δικαιώματα:

  • Κάνει εμφάνιση ειδοποιήσεων σε εσάς

Συλλογή δεδομένων:

  • Ο δημιουργός δηλώνει ότι αυτή η επέκταση δεν απαιτεί συλλογή δεδομένων.
Μάθετε περισσότερα
Περισσότερες πληροφορίες
Σύνδεσμοι προσθέτου
  • Email υποστήριξης
  • Αντιγραφή ID του πρόσθετου
Έκδοση
1.31
Μέγεθος
120,83 KB
Τελευταία ενημέρωση
4 μέρες πριν (1 Σεπ 2026)
Σχετικές κατηγορίες
  • Προγραμματισμός web
  • Διαχείριση λήψεων
  • Καρτέλες
Άδεια
Άδεια MIT
Ιστορικό εκδόσεων
  • Προβολή όλων των εκδόσεων
Ετικέτες
  • download
Προσθήκη σε συλλογή
Αναφορά προσθέτου
Μετάβαση στην αρχική σελίδα της Mozilla

Πρόσθετα

  • Σχετικά
  • Blog προσθέτων Firefox
  • Εργαστήριο επεκτάσεων
  • Κέντρο προγραμματιστών
  • Πολιτικές προγραμματιστών
  • Blog κοινότητας
  • Φόρουμ
  • Αναφορά σφάλματος
  • Οδηγίες κριτικής

Λήψη

  • Download Firefox
  • Windows
  • macOS
  • iOS
  • Android
  • Linux
  • All

Τελευταίες εκδόσεις

  • Nightly
  • Beta

Firefox για επιχειρήσεις

  • Enterprise

Κοινότητα

  • Connect
  • Contribute
  • Developer

Ακολουθήστε

  • Instagram
  • YouTube
  • TikTok
  • Bluesky
  • Podcast
  • Απόρρητο
  • Cookie
  • Νομικά

Εκτός από τα μέρη όπου αναφέρεται διαφορετικά, το περιεχόμενο του ιστοτόπου υπόκειται στην άδεια Creative Commons Attribution Share-Alike License v3.0 ή τυχόν νεότερες εκδόσεις. Το Android αποτελεί εμπορικό σήμα της Google LLC.