Skip to main content

Python MediaWiki Bot Framework

Project description

GitHub CI AppVeyor Build Status Code coverage Maintainability Python Top language Pywikibot release wheel Total downloads Monthly downloads Last commit

Pywikibot

The Pywikibot framework is a Python library that interfaces with the MediaWiki API version 1.27 or higher.

Also included are various general function scripts that can be adapted for different tasks.

For further information about the library excluding scripts see the full code documentation.

Quick start

pip install requests
git clone https://gerrit.wikimedia.org/r/pywikibot/core.git
cd core
git submodule update --init
python pwb.py script_name

Or to install using PyPI (excluding scripts)

pip install -U setuptools
pip install pywikibot
pwb <scriptname>

Our installation guide has more details for advanced usage.

Basic Usage

If you wish to write your own script it’s very easy to get started:

import pywikibot
site = pywikibot.Site('en', 'wikipedia')  # The site we want to run our bot on
page = pywikibot.Page(site, 'Wikipedia:Sandbox')
page.text = page.text.replace('foo', 'bar')
page.save('Replacing "foo" with "bar"')  # Saves the page

Wikibase Usage

Wikibase is a flexible knowledge base software that drives Wikidata. A sample pywikibot script for getting data from Wikibase:

import pywikibot
site = pywikibot.Site('wikipedia:en')
repo = site.data_repository()  # the Wikibase repository for given site
page = repo.page_from_repository('Q91')  # create a local page for the given item
item = pywikibot.ItemPage(repo, 'Q91')  # a repository item
data = item.get()  # get all item data from repository for this item

Script example

Pywikibot provides bot classes to develop your own script easily:

import pywikibot
from pywikibot import pagegenerators
from pywikibot.bot import ExistingPageBot

class MyBot(ExistingPageBot):

    update_options = {
        'text': 'This is a test text',
        'summary': 'Bot: a bot test edit with Pywikibot.'
    }

    def treat_page(self):
        """Load the given page, do some changes, and save it."""
        text = self.current_page.text
        text += '\n' + self.opt.text
        self.put_current(text, summary=self.opt.summary)

def main():
    """Parse command line arguments and invoke bot."""
    options = {}
    gen_factory = pagegenerators.GeneratorFactory()
    # Option parsing
    local_args = pywikibot.handle_args(args)  # global options
    local_args = gen_factory.handle_args(local_args)  # generators options
    for arg in local_args:
        opt, sep, value = arg.partition(':')
        if opt in ('-summary', '-text'):
            options[opt[1:]] = value
    MyBot(generator=gen_factory.getCombinedGenerator(), **options).run()

if __name == '__main__':
    main()

For more documentation on Pywikibot see our docs.

Roadmap

Current release

  • Add support for fonwiki (T347941)

  • site.BaseSite.redirects()and site.APISite.redirects() methods were added (T347226)

  • Upcast to pywikibot.FilePagefor a proper extension only (T346889)

  • Handle missing SDC mediainfo (T345038)

  • modules_only_mode parameter of data.api.ParamInfo, its paraminfo_keys class attribute and its preloaded_modules property was deprecated, the data.api.ParamInfo.normalize_paraminfo method became a staticmethod (T306637)

  • raise ValueError when pywikibot.FilePagetitle doesn’t have a valid file extension (T345786)

  • site.APISite.file_extensions property was added (T345786)

  • dropdelay and releasepid attributes of throttle.Throttlewhere deprecated in favour of expiry class attribute

  • Add https scheme if missing in url asked by pywikibot.scripts.generate_family_file

  • L10N updates and i18n updates

  • use inline re.IGNORECASE flag in textlib.case_escapefunction (T308265)

  • Convert URL-encoded characters also for links outside main namespace with cosmetic_changes.CosmeticChangesToolkit.cleanUpLinks(T342470)

  • Implement Flow topic summaries (T109443)

Deprecations

  • 8.4.0: Python 3.6 support is deprecated and will be dropped soon with Pywikibot 9

  • 8.4.0: modules_only_mode parameter of data.api.ParamInfo, its paraminfo_keys class attribute and its preloaded_modules property will be removed

  • 8.4.0: dropdelay and releasepid attributes of throttle.Throttlewill be removed in favour of expiry class attribute

  • 8.2.0: tools.itertools.itergroupwill be removed in favour of backports.batched

  • 8.2.0: normalize parameter of WbTime.toTimestrand WbTime.toWikibasewill be removed

  • 8.1.0: Dependency of exceptions.NoSiteLinkErrorfrom exceptions.NoPageErrorwill be removed

  • 8.1.0: exceptions.Server414Error is deprecated in favour of exceptions.Client414Error

  • 8.0.0: Timestamp.clone()method is deprecated in favour of Timestamp.replace() method.

  • 8.0.0: family.Family.maximum_GET_lengthmethod is deprecated in favour of config.maximum_GET_length(T325957)

  • 8.0.0: addOnly parameter of textlib.replaceLanguageLinksand textlib.replaceCategoryLinksare deprecated in favour of add_only

  • 8.0.0: textlib.TimeStripperregex attributes ptimeR, ptimeznR, pyearR, pmonthR, pdayR are deprecated in favour of patterns attribute which is a textlib.TimeStripperPatterns.

  • 8.0.0: textlib.TimeStripper``groups`` attribute is deprecated in favour of textlib.TIMEGROUPS

  • 8.0.0: LoginManager.get_login_tokenwas replaced by login.ClientLoginManager.site.tokens['login']

  • 8.0.0: data.api.LoginManager() is deprecated in favour of login.ClientLoginManager

  • 8.0.0: APISite.messages()method is deprecated in favour of userinfo[‘messages’]

  • 8.0.0: Page.editTime()method is deprecated and should be replaced by Page.latest_revision.timestamp

  • 7.7.0: tools.threadingclasses should no longer imported from tools

  • 7.6.0: tools.itertoolsdatatypes should no longer imported from tools

  • 7.6.0: tools.collectionsdatatypes should no longer imported from tools

  • 7.5.0: textlib.tzoneFixedOffset class will be removed in favour of time.TZoneFixedOffset

  • 7.4.0: FilePage.usingPages() was renamed to using_pages()

  • 7.2.0: tb parameter of exception()function was renamed to exc_info

  • 7.2.0: XMLDumpOldPageGenerator is deprecated in favour of a content parameter of XMLDumpPageGenerator(T306134)

  • 7.2.0: RedirectPageBot and NoRedirectPageBot bot classes are deprecated in favour of use_redirectsattribute

  • 7.2.0: tools.formatter.color_formatis deprecated and will be removed

  • 7.1.0: Unused get_redirect parameter of Page.getOldVersion()will be removed

  • 7.0.0: User.isBlocked() method is renamed to is_blocked for consistency

  • 7.0.0: A boolean watch parameter in Page.save() is deprecated and will be desupported

  • 7.0.0: baserevid parameter of editSource(), editQualifier(), removeClaims(), removeSources(), remove_qualifiers() DataSite methods will be removed

  • 7.0.0: Values of APISite.allpages() parameter filterredir other than True, False and None are deprecated

  • 7.0.0: The i18n identifier ‘cosmetic_changes-append’ will be removed in favour of ‘pywikibot-cosmetic-changes’

Will be removed in Pywikibot 9
  • 6.5.0: OutputOption.output() method will be removed in favour of OutputOption.out property

  • 6.5.0: Infinite rotating file handler with logfilecount of -1 is deprecated

  • 6.4.0: ‘allow_duplicates’ parameter of tools.itertools.intersect_generatorsas positional argument is deprecated, use keyword argument instead

  • 6.4.0: ‘iterables’ of tools.itertools.intersect_generatorsgiven as a list or tuple is deprecated, either use consecutive iterables or use ‘*’ to unpack

  • 6.2.0: outputter of OutputProxyOption without out property is deprecated

  • 6.2.0: ContextOption.output_range() and HighlightContextOption.output_range() are deprecated

  • 6.2.0: Error messages with ‘%’ style is deprecated in favour for str.format() style

  • 6.2.0: page.url2unicode() function is deprecated in favour of tools.chars.url2string()

  • 6.2.0: Throttle.multiplydelay attribute is deprecated

  • 6.2.0: SequenceOutputter.format_list() is deprecated in favour of ‘out’ property

  • 6.0.0: config.register_family_file() is deprecated

Release history

See https://github.com/wikimedia/pywikibot/blob/stable/HISTORY.rst

Contributing

Our code is maintained on Wikimedia’s Gerrit installation, learn how to get started.

Code of Conduct

The development of this software is covered by a Code of Conduct.

Project details


Release history Release notifications | RSS feed

This version

8.4.0

Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

pywikibot-8.4.0.tar.gz (605.0 kB view details)

Uploaded Source

Built Distribution

pywikibot-8.4.0-py3-none-any.whl (704.9 kB view details)

Uploaded Python 3

File details

Details for the file pywikibot-8.4.0.tar.gz.

File metadata

  • Download URL: pywikibot-8.4.0.tar.gz
  • Upload date:
  • Size: 605.0 kB
  • Tags: Source
  • Uploaded using Trusted Publishing? No
  • Uploaded via: twine/4.0.2 CPython/3.11.0

File hashes

Hashes for pywikibot-8.4.0.tar.gz
Algorithm Hash digest
SHA256 b919280ff9079742e01db38b6aabad826d65f5e5cc4d0949899d0b4510128f9d
MD5 aa33cd37e182029770e524fe1bb807b2
BLAKE2b-256 7db6f8132456f88a67b5912978bbb85e14d729d03af8ea0cf1f27f09d98deecc

See more details on using hashes here.

File details

Details for the file pywikibot-8.4.0-py3-none-any.whl.

File metadata

  • Download URL: pywikibot-8.4.0-py3-none-any.whl
  • Upload date:
  • Size: 704.9 kB
  • Tags: Python 3
  • Uploaded using Trusted Publishing? No
  • Uploaded via: twine/4.0.2 CPython/3.11.0

File hashes

Hashes for pywikibot-8.4.0-py3-none-any.whl
Algorithm Hash digest
SHA256 71ac28bd941f3f9b23d9449d4b165a25d72bc41b02d881253fd795b1e56dfed7
MD5 c82459aba581dfc8ba45b53814123422
BLAKE2b-256 43ddd49b93086c531e0d5f90596b39c818dbd43818b96b384fbc49889ffd5ca6

See more details on using hashes here.

Supported by

AWS AWS Cloud computing and Security Sponsor Datadog Datadog Monitoring Fastly Fastly CDN Google Google Download Analytics Microsoft Microsoft PSF Sponsor Pingdom Pingdom Monitoring Sentry Sentry Error logging StatusPage StatusPage Status page