E

8

164

I am using python Requests. I need to debug some OAuth activity, and for that I would like it to log all requests being performed. I could get this information with ngrep, but unfortunately it is not possible to grep https connections (which are needed for OAuth)

How can I activate logging of all URLs (+ parameters) that Requests is accessing?

Everhart answered 2/5, 2013 at 11:57 Comment(1)

The response by @yohann shows how to get yet more logging output, including the headers you're sending. It should be the accepted answer rather than Martijn's, which doesn't show the headers that you ended up getting via wireshark and hand-customizing a request instead. – Bunns 2/7, 2015 at 14:11

B

160

The underlying urllib3 library logs all new connections and URLs with the logging module, but not POST bodies. For GET requests this should be enough:

import logging

logging.basicConfig(level=logging.DEBUG)

which gives you the most verbose logging option; see the logging HOWTO for more details on how to configure logging levels and destinations.

Short demo:

>>> import requests
>>> import logging
>>> logging.basicConfig(level=logging.DEBUG)
>>> r = requests.get('http://httpbin.org/get?foo=bar&baz=python')
DEBUG:urllib3.connectionpool:Starting new HTTP connection (1): httpbin.org:80
DEBUG:urllib3.connectionpool:http://httpbin.org:80 "GET /get?foo=bar&baz=python HTTP/1.1" 200 366

Depending on the exact version of urllib3, the following messages are logged:

INFO: Redirects
WARN: Connection pool full (if this happens often increase the connection pool size)
WARN: Failed to parse headers (response headers with invalid format)
WARN: Retrying the connection
WARN: Certificate did not match expected hostname
WARN: Received response with both Content-Length and Transfer-Encoding, when processing a chunked response
DEBUG: New connections (HTTP or HTTPS)
DEBUG: Dropped connections
DEBUG: Connection details: method, path, HTTP version, status code and response length
DEBUG: Retry count increments

This doesn't include headers or bodies. urllib3 uses the http.client.HTTPConnection class to do the grunt-work, but that class doesn't support logging, it can normally only be configured to print to stdout. However, you can rig it to send all debug information to logging instead by introducing an alternative print name into that module:

import logging
import http.client

httpclient_logger = logging.getLogger("http.client")

def httpclient_logging_patch(level=logging.DEBUG):
    """Enable HTTPConnection debug logging to the logging framework"""

    def httpclient_log(*args):
        httpclient_logger.log(level, " ".join(args))

    # mask the print() built-in in the http.client module to use
    # logging instead
    http.client.print = httpclient_log
    # enable debugging
    http.client.HTTPConnection.debuglevel = 1

Calling httpclient_logging_patch() causes http.client connections to output all debug information to a standard logger, and so are picked up by logging.basicConfig():

>>> httpclient_logging_patch()
>>> r = requests.get('http://httpbin.org/get?foo=bar&baz=python')
DEBUG:urllib3.connectionpool:Starting new HTTP connection (1): httpbin.org:80
DEBUG:http.client:send: b'GET /get?foo=bar&baz=python HTTP/1.1\r\nHost: httpbin.org\r\nUser-Agent: python-requests/2.22.0\r\nAccept-Encoding: gzip, deflate\r\nAccept: */*\r\nConnection: keep-alive\r\n\r\n'
DEBUG:http.client:reply: 'HTTP/1.1 200 OK\r\n'
DEBUG:http.client:header: Date: Tue, 04 Feb 2020 13:36:53 GMT
DEBUG:http.client:header: Content-Type: application/json
DEBUG:http.client:header: Content-Length: 366
DEBUG:http.client:header: Connection: keep-alive
DEBUG:http.client:header: Server: gunicorn/19.9.0
DEBUG:http.client:header: Access-Control-Allow-Origin: *
DEBUG:http.client:header: Access-Control-Allow-Credentials: true
DEBUG:urllib3.connectionpool:http://httpbin.org:80 "GET /get?foo=bar&baz=python HTTP/1.1" 200 366

Bitner answered 2/5, 2013 at 12:5 Comment(12)

Strangely enough, I do not see the access_token in the OAuth request. Linkedin is complaining about unauthorized request, and I want to verify whether the library that I am using (rauth on top of requests) is sending that token with the request. I was expecting to see that as a query parameter, but maybe it is in the request headers? How can I force the urllib3 to show the headers too? And the request body? Just to make it simple: how can I see the FULL request? – Everhart 2/5, 2013 at 13:25

You cannot do that without patching, I'm afraid. The most common way to diagnose such problems is with a proxy or packet logger (I use wireshark to capture full requests and responses myself). I see you asked a new question on the subject though. – Bitner 2/5, 2013 at 14:6

Sure, I am debugging right now with wireshark, but I have a problem: if I do http, I see the full packet contents, but Linkedin returns 401, which is expected, since Linkedin tells to use https. But with https it is not working either, and I can not debug it since I can not inspect the TLS layer with wireshark. – Everhart 2/5, 2013 at 14:22

@gonvaled: Right, I see the problem. Then you have to patch requests to log headers (I'd expect to see Authorization headers for OAuth requests). – Bitner 2/5, 2013 at 14:33

Thanks @Martijn Pieters: I found the problem the following way: disable https, log request with wireshark, check that Linkedin replies with "Bearer header unkown", remove that header with bearer_auth=False for rauth, and set a query parameter "oauth2_access_token", send request again, log and see that linkedin is complaining now about "https required". Now moving to https solves the problem. I got lucky that Linkedin verifies first header and then http/https, otherwise I would have had no clue. I still think that logging full-packets is an intersting feature for urrlib3. I'll suggest this. – Everhart 2/5, 2013 at 14:45

@gonvaled: I was just looking into this, will post an answer to your other question shortly. – Bitner 2/5, 2013 at 14:46

You can get the headers that were sent without resorting to wireshark. See @yohann's answer. – Bunns 2/7, 2015 at 14:12

@nealmcb: gah, yes, setting a global class attribute would indeed enable debugging in httplib. I do wish that library used logging instead; the debug output is written directly to stdout rather than let you redirect it to a log destination of your choice. – Bitner 2/7, 2015 at 14:19

Hello, i'm using the logging module for different loging events. how can i disable the logging of the requests module? as implied above if you import both they run automatic. – Entomo 11/8, 2016 at 12:29

@TD_Nijboer: see How do I disable log messages from the Requests library? – Bitner 11/8, 2016 at 13:46

What does the 366 mean here? – Buttonhole 13/7, 2020 at 8:54

@Buttonhole the size of the response body, in bytes. Further down you can see DEBUG:http.client:header: Content-Length: 366. – Bitner 13/7, 2020 at 8:56

A

173

You need to enable debugging at httplib level (requests → urllib3 → httplib).

Here's some functions to both toggle (..._on() and ..._off()) or temporarily have it on:

import logging
import contextlib
from http.client import HTTPConnection

def debug_requests_on():
    '''Switches on logging of the requests module.'''
    HTTPConnection.debuglevel = 1

    logging.basicConfig()
    logging.getLogger().setLevel(logging.DEBUG)
    requests_log = logging.getLogger("requests.packages.urllib3")
    requests_log.setLevel(logging.DEBUG)
    requests_log.propagate = True

def debug_requests_off():
    '''Switches off logging of the requests module, might be some side-effects'''
    HTTPConnection.debuglevel = 0

    root_logger = logging.getLogger()
    root_logger.setLevel(logging.WARNING)
    root_logger.handlers = []
    requests_log = logging.getLogger("requests.packages.urllib3")
    requests_log.setLevel(logging.WARNING)
    requests_log.propagate = False

@contextlib.contextmanager
def debug_requests():
    '''Use with 'with'!'''
    debug_requests_on()
    yield
    debug_requests_off()

Demo use:

>>> requests.get('http://httpbin.org/')
<Response [200]>

>>> debug_requests_on()
>>> requests.get('http://httpbin.org/')
INFO:requests.packages.urllib3.connectionpool:Starting new HTTP connection (1): httpbin.org
DEBUG:requests.packages.urllib3.connectionpool:"GET / HTTP/1.1" 200 12150
send: 'GET / HTTP/1.1\r\nHost: httpbin.org\r\nConnection: keep-alive\r\nAccept-
Encoding: gzip, deflate\r\nAccept: */*\r\nUser-Agent: python-requests/2.11.1\r\n\r\n'
reply: 'HTTP/1.1 200 OK\r\n'
header: Server: nginx
...
<Response [200]>

>>> debug_requests_off()
>>> requests.get('http://httpbin.org/')
<Response [200]>

>>> with debug_requests():
...     requests.get('http://httpbin.org/')
INFO:requests.packages.urllib3.connectionpool:Starting new HTTP connection (1): httpbin.org
...
<Response [200]>

You will see the REQUEST, including HEADERS and DATA, and RESPONSE with HEADERS but without DATA. The only thing missing will be the response.body which is not logged.

Source

Aenneea answered 5/7, 2014 at 16:16 Comment(11)

Thank you for the insight about using httplib.HTTPConnection.debuglevel = 1 to get the headers - excellent! But I think I get the same results using just logging.basicConfig(level=logging.DEBUG) in place of your other 5 lines. Am I missing something? I guess it could be a way to set different logging levels for the root vs the urllib3, if desired. – Bunns 2/7, 2015 at 15:22

You haven't the header with your solution. – Aenneea 29/7, 2015 at 15:44

httplib.HTTPConnection.debuglevel = 2 will allow printing of POST body as well. – Latricelatricia 30/10, 2015 at 9:58

httplib.HTTPConnection.debuglevel = 1 is enough @Latricelatricia $ curl https://raw.githubusercontent.com/python/cpython/master/Lib/http/client.py |grep debuglevel it's always debuglevel > 0 – Aenneea 27/4, 2016 at 10:52

You are a rockstar man. Im trying to debug requests to know why my cookies are empty on some sites, any hints? – Scolopendrid 25/12, 2018 at 13:1

This works great for request headers, but response headers are showing blank, eg. GET / HTTP/1.1" 200 None header: Content-Encoding header: Cache-Control header: ... – Ladew 19/2, 2019 at 21:46

Someway to prevent the logged content to be sent to the standard output ? – Oracle 19/8, 2019 at 14:7

It works only for stdout but not for log file. Problem example here: https://mcmap.net/q/151600/-python-http-request-and-debug-level-logging-to-the-log-file/1090360 – Feder 7/11, 2019 at 15:24

I faced a weird issue due to this - the fact that root logger's handlers are removed caused a bug on my production that was super hard to debug #72950015 – Revels 21/7, 2022 at 9:56

It is not printing the actual response – Marj 23/11, 2022 at 5:4

This is a great solution, thanks! I don't understand, @Yohann, why you remove the handlers with root_logger.handlers = []. You're not adding any handlers, so why remove them? I presume this is what you mean when you say there "might be some side-effects", but I don't understand why you even have that line. Otherwise, this is a great, almost fully generalized solution for any dependent library introspection. – France 4/4 at 8:27

B

160