Skip to main content

Unofficial Qwen API Client & Chat Wrapper

Project description

qwen-chat 🚀

PyPI version License: MIT Python Version

An unofficial, feature-rich Python SDK and client wrapper for the Qwen AI Web API. Enjoy seamless access to advanced models like qwen3.6-plus with support for streaming, multi-turn conversations, system prompts, web search, deep reasoning, and automatic local image uploads.


✨ Features

  • 🧠 Complete Model Support
    Interact with Qwen's latest web models including qwen3.6-plus, qwen-max-latest, qwq-32b, qwen2.5-omni-7b, and specialized vision/coder versions.

  • 💬 Multi-turn Chat Sessions
    Have back-and-forth conversations that maintain context across messages — just like chatting in the Qwen web UI. Build interactive terminal chatbots or conversational agents easily.

  • ⚡ Synchronous & Asynchronous Clients
    Whether you're building a script or a highly concurrent async web server, qwen-chat has you covered with native create() and acreate() workflows.

  • 🌊 Real-time SSE Streaming (On/Off)
    Toggle stream=True or stream=False per request. Stream token-by-token responses directly to your UI or console, or get the full response at once.

  • 🎭 System & Assistant Roles
    Use system role messages to define the AI's personality, behavior, and instructions. Chain assistant and user roles for few-shot prompting and context injection.

  • 📸 Automatic Local Image Uploads
    Pass a local file path or raw bytes to an ImageBlock. The client automatically fetches STS tokens and uploads the image to Alibaba Cloud OSS under the hood — no boilerplate required!

  • 🔍 Integrated Web Search
    Toggle real-time search on or off per message to query live web data and receive detailed citations alongside responses.

  • 💡 Thinking & Reasoning Controls
    Adjust thinking budget settings to enable advanced reasoning capabilities on complex tasks.


📦 Installation

Install qwen-chat via pip:

pip install qwen-chat

⚙️ Environment Setup

To authenticate requests, extract your Authorization bearer token from the Qwen Web UI:

  1. Go to https://chat.qwen.ai and log in.
  2. Open developer tools (F12 or Ctrl+Shift+I / Cmd+Option+I) and navigate to the Network tab.
  3. Send any message in the chat interface.
  4. Locate the completions request (filter by Fetch/XHR).
  5. Click on the request and go to the Headers tab. Copy the value of the Authorization header without the word "Bearer " (just copy the token starting with eyJ...).
  6. Save this value in a .env file in the root of your project:
QWEN_AUTH_TOKEN=eyJhbGciOiJIUzI1NiIsInR5cCI6IkpXVCJ9...

Note: You do not need to set up cookies. The package handles cookies automatically.


🚀 Usage Examples

1. Basic Text Completion

The simplest way to get a response from Qwen:

from qwen_chat import Qwen
from qwen_chat.core.types.chat import ChatMessage

client = Qwen()

messages = [
    ChatMessage(role="user", content="What is Python?")
]

response = client.chat.create(messages=messages, model="qwen3.6-plus")
print("🤖", response.choices.message.content)

2. Streaming On / Off

Stream OFF (get full response at once):

from qwen_chat import Qwen
from qwen_chat.core.types.chat import ChatMessage

client = Qwen()
messages = [ChatMessage(role="user", content="Tell me a short joke.")]

# stream=False (default) — returns complete response
response = client.chat.create(messages=messages, model="qwen3.6-plus", stream=False)
print("🤖", response.choices.message.content)

Stream ON (token-by-token output):

from qwen_chat import Qwen
from qwen_chat.core.types.chat import ChatMessage

client = Qwen()
messages = [ChatMessage(role="user", content="Write a poem about the ocean.")]

# stream=True — yields chunks in real-time
stream = client.chat.create(messages=messages, model="qwen3.6-plus", stream=True)

print("🤖 ", end="")
for chunk in stream:
    print(chunk.choices[0].delta.content, end="", flush=True)
print()

3. System Role & Assistant Role (Custom Personality)

Use the system role to define how the AI behaves. Use assistant role for few-shot examples:

from qwen_chat import Qwen
from qwen_chat.core.types.chat import ChatMessage

client = Qwen()

messages = [
    # System prompt — defines AI personality and behavior
    ChatMessage(
        role="system",
        content="You are a friendly pirate captain. Always respond in pirate speak with lots of 'Arrr!' and nautical references."
    ),
    # User message
    ChatMessage(
        role="user",
        content="What's the weather like today?"
    )
]

response = client.chat.create(messages=messages, model="qwen3.6-plus")
print("🏴‍☠️", response.choices.message.content)

Few-shot prompting with assistant role:

from qwen_chat import Qwen
from qwen_chat.core.types.chat import ChatMessage

client = Qwen()

messages = [
    ChatMessage(role="system", content="You are a helpful translator. Translate English to Bengali."),
    # Few-shot example
    ChatMessage(role="user", content="Hello"),
    ChatMessage(role="assistant", content="হ্যালো"),
    ChatMessage(role="user", content="How are you?"),
    ChatMessage(role="assistant", content="আপনি কেমন আছেন?"),
    # Actual request
    ChatMessage(role="user", content="I love programming")
]

response = client.chat.create(messages=messages, model="qwen3.6-plus")
print("🤖", response.choices.message.content)

4. Interactive Terminal Chat (Multi-turn Conversation)

Build a fully interactive terminal chatbot that maintains conversation context across multiple turns — just like the Qwen web UI:

from qwen_chat import Qwen
from qwen_chat.core.types.chat import ChatMessage

def main():
    client = Qwen()
    print("💬 Qwen Chat Terminal")
    print("Type 'exit' to quit\n")

    history = [
        ChatMessage(
            role="system",
            content="You are a helpful and friendly AI assistant."
        )
    ]

    while True:
        try:
            user_input = input("🧑 You: ").strip()
        except (EOFError, KeyboardInterrupt):
            print("\nBye! 👋")
            break

        if user_input.lower() == "exit":
            print("Bye! 👋")
            break

        # Add user message to conversation history
        history.append(ChatMessage(role="user", content=user_input))

        # Send full history for context-aware response
        response = client.chat.create(
            messages=history,
            model="qwen3.6-plus",
            temperature=0.7
        )

        reply = response.choices.message.content
        print(f"🤖 AI: {reply}\n")

        # Add assistant reply to history for next turn
        history.append(ChatMessage(role="assistant", content=reply))

if __name__ == "__main__":
    main()

Interactive chat with streaming output:

from qwen_chat import Qwen
from qwen_chat.core.types.chat import ChatMessage

def main():
    client = Qwen()
    print("💬 Qwen Streaming Chat")
    print("Type 'exit' to quit\n")

    history = []

    while True:
        try:
            user_input = input("🧑 You: ").strip()
        except (EOFError, KeyboardInterrupt):
            print("\nBye! 👋")
            break

        if user_input.lower() == "exit":
            print("Bye! 👋")
            break

        history.append(ChatMessage(role="user", content=user_input))

        # Stream the response token-by-token
        stream = client.chat.create(
            messages=history,
            model="qwen3.6-plus",
            stream=True
        )

        print("🤖 AI: ", end="")
        full_reply = ""
        for chunk in stream:
            text = chunk.choices[0].delta.content
            print(text, end="", flush=True)
            full_reply += text
        print("\n")

        history.append(ChatMessage(role="assistant", content=full_reply))

if __name__ == "__main__":
    main()

5. Web Search with Citations

from qwen_chat import Qwen
from qwen_chat.core.types.chat import ChatMessage

client = Qwen()

messages = [
    ChatMessage(
        role="user",
        content="What are the latest developments in AI?",
        web_search=True
    )
]

stream = client.chat.create(messages=messages, model="qwen3.6-plus", stream=True)

for chunk in stream:
    delta = chunk.choices[0].delta
    if delta.extra and delta.extra.web_search_info:
        print("\n🔍 Sources:")
        for source in delta.extra.web_search_info:
            print(f"  • {source.title}: {source.url}")
        print()
    print(delta.content, end="", flush=True)
print()

6. Automatic Local Image Upload (Vision)

Just pass your local image path directly inside ImageBlock. The client handles the secure upload automatically:

from qwen_chat import Qwen
from qwen_chat.core.types.chat import ChatMessage, TextBlock, ImageBlock

client = Qwen()

messages = [
    ChatMessage(
        role="user",
        blocks=[
            TextBlock(text="What's in this image? Describe it in detail."),
            ImageBlock(path="photo.jpg")  # Auto-uploaded to Alibaba OSS!
        ]
    )
]

response = client.chat.create(messages=messages, model="qwen3.6-plus")
print("🤖", response.choices.message.content)

7. Async Usage

import asyncio
from qwen_chat import Qwen
from qwen_chat.core.types.chat import ChatMessage

async def main():
    client = Qwen()
    messages = [ChatMessage(role="user", content="Explain quantum computing in one sentence.")]
    
    response = await client.chat.acreate(messages=messages, model="qwen3.6-plus")
    print("🤖", response.choices.message.content)

asyncio.run(main())

8. Thinking & Reasoning Mode

Enable deep thinking for complex tasks:

from qwen_chat import Qwen
from qwen_chat.core.types.chat import ChatMessage

client = Qwen()

messages = [
    ChatMessage(
        role="user",
        content="Solve this step by step: If a train travels 120km in 2 hours, then stops for 30 minutes, then travels 90km in 1.5 hours, what is the average speed for the entire journey?",
        thinking=True,
        thinking_budget=4096
    )
]

response = client.chat.create(messages=messages, model="qwq-32b")
print("🤖", response.choices.message.content)

🙋‍♂️ Contributing

Contributions are welcome! Here's how:

  1. Fork the project
  2. Create your feature branch (git checkout -b feature/awesome-feature)
  3. Commit your changes (git commit -m 'Add awesome feature')
  4. Push to the branch (git push origin feature/awesome-feature)
  5. Open a Pull Request

📃 License

This project is licensed under the MIT License.


📞 Contact & Support

Telegram

For queries, support, or custom integrations, feel free to reach out!


Made with ❤️ by Shahadat Hassan

Project details


Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

qwen_chat-1.0.2.tar.gz (34.7 kB view details)

Uploaded Source

Built Distribution

If you're not sure about the file name format, learn more about wheel file names.

qwen_chat-1.0.2-py3-none-any.whl (35.2 kB view details)

Uploaded Python 3

File details

Details for the file qwen_chat-1.0.2.tar.gz.

File metadata

  • Download URL: qwen_chat-1.0.2.tar.gz
  • Upload date:
  • Size: 34.7 kB
  • Tags: Source
  • Uploaded using Trusted Publishing? No
  • Uploaded via: twine/6.2.0 CPython/3.12.7

File hashes

Hashes for qwen_chat-1.0.2.tar.gz
Algorithm Hash digest
SHA256 fe018d94e6b58bbe78bcb03d1f9c63f6ff349886107230293b816d3837fd73ce
MD5 0ca6f416ac73e3d9171ad755bc876a18
BLAKE2b-256 de68c9bf9ed12809af5072144eb64aec8dd99cb1f677a8e2837a05afcdf23b1e

See more details on using hashes here.

File details

Details for the file qwen_chat-1.0.2-py3-none-any.whl.

File metadata

  • Download URL: qwen_chat-1.0.2-py3-none-any.whl
  • Upload date:
  • Size: 35.2 kB
  • Tags: Python 3
  • Uploaded using Trusted Publishing? No
  • Uploaded via: twine/6.2.0 CPython/3.12.7

File hashes

Hashes for qwen_chat-1.0.2-py3-none-any.whl
Algorithm Hash digest
SHA256 25ed4f0f56f404fb091bc8386a589d223770934496d655587721ff09e6342415
MD5 3341d66dc154b57fca060e377a8623e5
BLAKE2b-256 d1567c878b6f8195ab1f7b4c28299a4e9ad84beec2e6676a2adaf61c55aacea1

See more details on using hashes here.

Supported by

AWS Cloud computing and Security Sponsor Datadog Monitoring Depot Continuous Integration Fastly CDN Google Download Analytics Pingdom Monitoring Sentry Error logging StatusPage Status page