← eval3_submission_v2

requests_7427

resolved RESOLVED PASS · 12 tool calls · 164 s · psf/requests

Task input

Port bpo-39057 to Requests

This PR ports a [fix](https://bugs.python.org/issue39057) from CPython to constrain no_proxy behavior to domain boundaries. This was originally undefined behavior when Requests implemented `no_proxy` support but has since had bug fixes released in both CPython and Curl. We'll follow suit to conform with ecosystem expectations and resolve long standing issue #4795.

Tool calls (13)

#ToolArgumentsResult
1run_command
{
 "command": "grep -rn \"no_proxy\" src/requests | head -20"
}
{"status": "ok", "stdout": "src/requests/sessions.py:840: no_proxy = proxies.get(\"no_proxy\") if proxies is not None else None\nsrc/requests/sessions.py:841: e
full
{"status": "ok", "stdout": "src/requests/sessions.py:840:            no_proxy = proxies.get(\"no_proxy\") if proxies is not None else None\nsrc/requests/sessions.py:841:            env_proxies = get_environ_proxies(url, no_proxy=no_proxy)\nsrc/requests/utils.py:765:    Very simple check of the cidr format in no_proxy variable.\nsrc/requests/utils.py:810:def should_bypass_proxies(url: str, no_proxy: str | None) -> bool:\nsrc/requests/utils.py:822:    # First check whether no_proxy is defined. If it is, check that the URL\nsrc/requests/utils.py:823:    # we're getting isn't in the no_proxy list.\nsrc/requests/utils.py:824:    no_proxy_arg = no_proxy\nsrc/requests/utils.py:825:    if no_proxy is None:\nsrc/requests/utils.py:826:        no_proxy = get_proxy(\"no_proxy\")\nsrc/requests/utils.py:834:    if no_proxy:\nsrc/requests/utils.py:837:        no_proxy_hosts = (host for host in no_proxy.replace(\" \", \"\").split(\",\") if host)\nsrc/requests/utils.py:840:            for proxy_ip in no_proxy_hosts:\nsrc/requests/utils.py:845:                    # If no_proxy ip was defined in plain IP notation instead of cidr notation &\nsrc/requests/utils.py:853:            for host in no_proxy_hosts:\nsrc/requests/utils.py:855:                    # The URL does match something in no_proxy, so we don't want\nsrc/requests/utils.py:859:    with set_environ(\"no_proxy\", no_proxy_arg):\nsrc/requests/utils.py:871:def get_environ_proxies(url: str, no_proxy: str | None = None) -> dict[str, str]:\nsrc/requests/utils.py:877:    if should_bypass_proxies(url, no_proxy=no_proxy):\nsrc/requests/utils.py:927:    no_proxy = proxies.get(\"no_proxy\")\nsrc/requests/utils.py:930:    if trust_env and not should_bypass_proxies(url, no_proxy=no_proxy):\n", "exit_code": 0}
2read_file
{
 "filepath": "src/requests/utils.py",
 "start_line": 810,
 "end_line": 865
}
{"status": "ok", "filepath": "src/requests/utils.py", "content": "def should_bypass_proxies(url: str, no_proxy: str | None) -> bool:\n \"\"\"\n Returns whether
full
{"status": "ok", "filepath": "src/requests/utils.py", "content": "def should_bypass_proxies(url: str, no_proxy: str | None) -> bool:\n    \"\"\"\n    Returns whether we should bypass proxies or not.\n\n    :rtype: bool\n    \"\"\"\n\n    # Prioritize lowercase environment variables over uppercase\n    # to keep a consistent behaviour with other http projects (curl, wget).\n    def get_proxy(key: str) -> str | None:\n        return os.environ.get(key) or os.environ.get(key.upper())\n\n    # First check whether no_proxy is defined. If it is, check that the URL\n    # we're getting isn't in the no_proxy list.\n    no_proxy_arg = no_proxy\n    if no_proxy is None:\n        no_proxy = get_proxy(\"no_proxy\")\n    parsed = urlparse(url)\n    hostname = parsed.hostname\n\n    if hostname is None:\n        # URLs don't always have hostnames, e.g. file:/// urls.\n        return True\n\n    if no_proxy:\n        # We need to check whether we match here. We need to see if we match\n        # the end of the hostname, both with and without the port.\n        no_proxy_hosts = (host for host in no_proxy.replace(\" \", \"\").split(\",\") if host)\n\n        if is_ipv4_address(hostname):\n            for proxy_ip in no_proxy_hosts:\n                if is_valid_cidr(proxy_ip):\n                    if address_in_network(hostname, proxy_ip):\n                        return True\n                elif hostname == proxy_ip:\n                    # If no_proxy ip was defined in plain IP notation instead of cidr notation &\n                    # matches the IP of the index\n                    return True\n        else:\n            host_with_port = hostname\n            if parsed.port:\n                host_with_port += f\":{parsed.port}\"\n\n            for host in no_proxy_hosts:\n                if hostname.endswith(host) or host_with_port.endswith(host):\n                    # The URL does match something in no_proxy, so we don't want\n                    # to apply the proxies on this 
3run_command
{
 "command": "python3 -c \"import requests; from requests.utils import should_bypass_proxies; print(should_bypass_proxies('http://my-api.example.com', 'api.example.com')); print(should_bypass_proxies('http://not-api.example.com', 'api.example.com'))\""
}
{"status": "ok", "stdout": "True\nTrue\n", "exit_code": 0}
full
{"status": "ok", "stdout": "True\nTrue\n", "exit_code": 0}
4read_file
{
 "end_line": 865,
 "filepath": "src/requests/utils.py",
 "start_line": 810
}
{"status": "ok", "filepath": "src/requests/utils.py", "content": "def should_bypass_proxies(url: str, no_proxy: str | None) -> bool:\n \"\"\"\n Returns whether
full
{"status": "ok", "filepath": "src/requests/utils.py", "content": "def should_bypass_proxies(url: str, no_proxy: str | None) -> bool:\n    \"\"\"\n    Returns whether we should bypass proxies or not.\n\n    :rtype: bool\n    \"\"\"\n\n    # Prioritize lowercase environment variables over uppercase\n    # to keep a consistent behaviour with other http projects (curl, wget).\n    def get_proxy(key: str) -> str | None:\n        return os.environ.get(key) or os.environ.get(key.upper())\n\n    # First check whether no_proxy is defined. If it is, check that the URL\n    # we're getting isn't in the no_proxy list.\n    no_proxy_arg = no_proxy\n    if no_proxy is None:\n        no_proxy = get_proxy(\"no_proxy\")\n    parsed = urlparse(url)\n    hostname = parsed.hostname\n\n    if hostname is None:\n        # URLs don't always have hostnames, e.g. file:/// urls.\n        return True\n\n    if no_proxy:\n        # We need to check whether we match here. We need to see if we match\n        # the end of the hostname, both with and without the port.\n        no_proxy_hosts = (host for host in no_proxy.replace(\" \", \"\").split(\",\") if host)\n\n        if is_ipv4_address(hostname):\n            for proxy_ip in no_proxy_hosts:\n                if is_valid_cidr(proxy_ip):\n                    if address_in_network(hostname, proxy_ip):\n                        return True\n                elif hostname == proxy_ip:\n                    # If no_proxy ip was defined in plain IP notation instead of cidr notation &\n                    # matches the IP of the index\n                    return True\n        else:\n            host_with_port = hostname\n            if parsed.port:\n                host_with_port += f\":{parsed.port}\"\n\n            for host in no_proxy_hosts:\n                if hostname.endswith(host) or host_with_port.endswith(host):\n                    # The URL does match something in no_proxy, so we don't want\n                    # to apply the proxies on this 
5edit_file
{
 "filepath": "src/requests/utils.py",
 "new_string": "            for host in no_proxy_hosts:\n                if hostname == host or hostname.endswith(\\\".\\\" + host) or host_with_port == host or host_with_port.endswith(\\\".\\\" + host):\n                    # The URL does match something in no_proxy, so we don't want\n                    # to apply the proxies on this URL.\n                    return True",
 "old_string": "            for host in no_proxy_hosts:\n                if hostname.endswith(host) or host_with_port.endswith(host):\n                    # The URL does match something in no_proxy, so we don't want\n                    # to apply the proxies on this URL.\n                    return True"
}
{"status": "ok", "filepath": "src/requests/utils.py", "occurrences": 1, "strategy": "exact", "diff": "--- a/src/requests/utils.py\n+++ b/src/requests/utils.py\n
full
{"status": "ok", "filepath": "src/requests/utils.py", "occurrences": 1, "strategy": "exact", "diff": "--- a/src/requests/utils.py\n+++ b/src/requests/utils.py\n@@ -851,7 +851,7 @@\n                 host_with_port += f\":{parsed.port}\"\n \n             for host in no_proxy_hosts:\n-                if hostname.endswith(host) or host_with_port.endswith(host):\n+                if hostname == host or hostname.endswith(\\\".\\\" + host) or host_with_port == host or host_with_port.endswith(\\\".\\\" + host):\n                     # The URL does match something in no_proxy, so we don't want\n                     # to apply the proxies on this URL.\n                     return True\n", "is_truncated": false}
6run_command
{
 "command": "python3 -c \"import requests; from requests.utils import should_bypass_proxies; print(f'Match: {should_bypass_proxies(\\'http://api.example.com\\', \\'api.example.com\\')}'); print(f'Subdomain: {should_bypass_proxies(\\'http://sub.api.example.com\\', \\'api.example.com\\')}'); print(f'Wrong domain: {should_bypass_proxies(\\'http://not-api.example.com\\', \\'api.example.com\\')}'); print(f'Port match: {should_bypass_proxies(\\'http://api.example.com:8080\\', \\'api.example.com:8080\\')}'); print(f'Port subdomain: {should_bypass_proxies(\\'http://sub.api.example.com:8080\\', \\'api.example.com:8080\\')}')\""
}
{"status": "error", "error_type": "CommandError", "error_message": "Command failed with exit code 1", "details": {"stdout": "", "stderr": " File \"<string>\", l
full
{"status": "error", "error_type": "CommandError", "error_message": "Command failed with exit code 1", "details": {"stdout": "", "stderr": "  File \"<string>\", line 1\n    import requests; from requests.utils import should_bypass_proxies; print(f'Match: {should_bypass_proxies(\\'http://api.example.com\\', \\'api.example.com\\')}'); print(f'Subdomain: {should_bypass_proxies(\\'http://sub.api.example.com\\', \\'api.example.com\\')}'); print(f'Wrong domain: {should_bypass_proxies(\\'http://not-api.example.com\\', \\'api.example.com\\')}'); print(f'Port match: {should_bypass_proxies(\\'http://api.example.com:8080\\', \\'api.example.com:8080\\')}'); print(f'Port subdomain: {should_bypass_proxies(\\'http://sub.api.example.com:8080\\', \\'api.example.com:8080\\')}')\n                                                                                                              ^\nSyntaxError: unexpected character after line continuation character\n", "exit_code": 1}}
7run_command
{
 "command": "python3 -c \"from requests.utils import should_bypass_proxies; print(f'Match: {should_bypass_proxies(\\\"http://api.example.com\\\", \\\"api.example.com\\\")}'); print(f'Subdomain: {should_bypass_proxies(\\\"http://sub.api.example.com\\\", \\\"api.example.com\\\")}'); print(f'Wrong domain: {should_bypass_proxies(\\\"http://not-api.example.com\\\", \\\"api.example.com\\\")}'); print(f'Port match: {should_bypass_proxies(\\\"http://api.example.com:8080\\\", \\\"api.example.com:8080\\\")}'); print(f'Port subdomain: {should_bypass_proxies(\\\"http://sub.api.example.com:8080\\\", \\\"api.example.com:8080\\\")}')\""
}
{"status": "error", "error_type": "CommandError", "error_message": "Command failed with exit code 1", "details": {"stdout": "", "stderr": "Traceback (most recen
full
{"status": "error", "error_type": "CommandError", "error_message": "Command failed with exit code 1", "details": {"stdout": "", "stderr": "Traceback (most recent call last):\n  File \"<string>\", line 1, in <module>\n    from requests.utils import should_bypass_proxies; print(f'Match: {should_bypass_proxies(\"http://api.example.com\", \"api.example.com\")}'); print(f'Subdomain: {should_bypass_proxies(\"http://sub.api.example.com\", \"api.example.com\")}'); print(f'Wrong domain: {should_bypass_proxies(\"http://not-api.example.com\", \"api.example.com\")}'); print(f'Port match: {should_bypass_proxies(\"http://api.example.com:8080\", \"api.example.com:8080\")}'); print(f'Port subdomain: {should_bypass_proxies(\"http://sub.api.example.com:8080\", \"api.example.com:8080\")}')\n    ^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^\n  File \"/workspace/src/requests/__init__.py\", line 158, in <module>\n    from . import packages, utils\n  File \"/workspace/src/requests/utils.py\", line 854\n    if hostname == host or hostname.endswith(\\\".\\\" + host) or host_with_port == host or host_with_port.endswith(\\\".\\\" + host):\n                                              ^\nSyntaxError: unexpected character after line continuation character\n", "exit_code": 1}}
8read_file
{
 "end_line": 860,
 "filepath": "src/requests/utils.py",
 "start_line": 845
}
{"status": "ok", "filepath": "src/requests/utils.py", "content": " # If no_proxy ip was defined in plain IP notation instead of cidr notation &\n # matches the
full
{"status": "ok", "filepath": "src/requests/utils.py", "content": "                    # If no_proxy ip was defined in plain IP notation instead of cidr notation &\n                    # matches the IP of the index\n                    return True\n        else:\n            host_with_port = hostname\n            if parsed.port:\n                host_with_port += f\":{parsed.port}\"\n\n            for host in no_proxy_hosts:\n                if hostname == host or hostname.endswith(\\\".\\\" + host) or host_with_port == host or host_with_port.endswith(\\\".\\\" + host):\n                    # The URL does match something in no_proxy, so we don't want\n                    # to apply the proxies on this URL.\n                    return True\n\n    with set_environ(\"no_proxy\", no_proxy_arg):\n        try:\n", "start_line": 845, "end_line": 860, "total_lines": 1153, "is_truncated": false}
9edit_file
{
 "filepath": "src/requests/utils.py",
 "new_string": "            for host in no_proxy_hosts:\n                if hostname == host or hostname.endswith(\".\" + host) or host_with_port == host or host_with_port.endswith(\".\" + host):\n                    # The URL does match something in no_proxy, so we don't want\n                    # to apply the proxies on this URL.\n                    return True",
 "old_string": "            for host in no_proxy_hosts:\n                if hostname == host or hostname.endswith(\\\\\\\\\\\\\\\".\\\\\\\\\\\\\\\" + host) or host_with_port == host or host_with_port.endswith(\\\\\\\\\\\\\\\".\\\\\\\\\\\\\\\" + host):\n                    # The URL does match something in no_proxy, so we don't want\n                    # to apply the proxies on this URL.\n                    return True"
}
{"status": "error", "error_type": "FileEditError", "error_message": "old_string not found in file (tried exact, flexible and regex matching)"}
full
{"status": "error", "error_type": "FileEditError", "error_message": "old_string not found in file (tried exact, flexible and regex matching)"}
10read_file
{
 "end_line": 858,
 "filepath": "src/requests/utils.py",
 "start_line": 850
}
{"status": "ok", "filepath": "src/requests/utils.py", "content": " if parsed.port:\n host_with_port += f\":{parsed.port}\"\n\n for host in no_proxy_hosts:\n if
full
{"status": "ok", "filepath": "src/requests/utils.py", "content": "            if parsed.port:\n                host_with_port += f\":{parsed.port}\"\n\n            for host in no_proxy_hosts:\n                if hostname == host or hostname.endswith(\\\".\\\" + host) or host_with_port == host or host_with_port.endswith(\\\".\\\" + host):\n                    # The URL does match something in no_proxy, so we don't want\n                    # to apply the proxies on this URL.\n                    return True\n\n", "start_line": 850, "end_line": 858, "total_lines": 1153, "is_truncated": false}
11edit_file
{
 "filepath": "src/requests/utils.py",
 "new_string": "            for host in no_proxy_hosts:\n                if hostname == host or hostname.endswith(\".\" + host) or host_with_port == host or host_with_port.endswith(\".\" + host):\n                    # The URL does match something in no_proxy, so we don't want\n                    # to apply the proxies on this URL.\n                    return True",
 "old_string": "            for host in no_proxy_hosts:\n                if hostname == host or hostname.endswith(\\\".\\\" + host) or host_with_port == host or host_with_port.endswith(\\\".\\\" + host):\n                    # The URL does match something in no_proxy, so we don't want\n                    # to apply the proxies on this URL.\n                    return True",
 "allow_multiple": true
}
{"status": "ok", "filepath": "src/requests/utils.py", "occurrences": 1, "strategy": "exact", "diff": "--- a/src/requests/utils.py\n+++ b/src/requests/utils.py\n
full
{"status": "ok", "filepath": "src/requests/utils.py", "occurrences": 1, "strategy": "exact", "diff": "--- a/src/requests/utils.py\n+++ b/src/requests/utils.py\n@@ -851,7 +851,7 @@\n                 host_with_port += f\":{parsed.port}\"\n \n             for host in no_proxy_hosts:\n-                if hostname == host or hostname.endswith(\\\".\\\" + host) or host_with_port == host or host_with_port.endswith(\\\".\\\" + host):\n+                if hostname == host or hostname.endswith(\".\" + host) or host_with_port == host or host_with_port.endswith(\".\" + host):\n                     # The URL does match something in no_proxy, so we don't want\n                     # to apply the proxies on this URL.\n                     return True\n", "is_truncated": false}
12run_command
{
 "command": "python3 -c \"from requests.utils import should_bypass_proxies; print(f'Match: {should_bypass_proxies(\\\"http://api.example.com\\\", \\\"api.example.com\\\")}'); print(f'Subdomain: {should_bypass_proxies(\\\"http://sub.api.example.com\\\", \\\"api.example.com\\\")}'); print(f'Wrong domain: {should_bypass_proxies(\\\"http://not-api.example.com\\\", \\\"api.example.com\\\")}'); print(f'Port match: {should_bypass_proxies(\\\"http://api.example.com:8080\\\", \\\"api.example.com:8080\\\")}'); print(f'Port subdomain: {should_bypass_proxies(\\\"http://sub.api.example.com:8080\\\", \\\"api.example.com:8080\\\")}')\""
}
{"status": "ok", "stdout": "Match: True\nSubdomain: True\nWrong domain: False\nPort match: True\nPort subdomain: True\n", "exit_code": 0}
full
{"status": "ok", "stdout": "Match: True\nSubdomain: True\nWrong domain: False\nPort match: True\nPort subdomain: True\n", "exit_code": 0}
13submit_patch
[]
{"status": "ok", "patch_size": 716, "files_changed": 1}
full
{"status": "ok", "patch_size": 716, "files_changed": 1}

Patch

diff --git a/src/requests/utils.py b/src/requests/utils.py
index e69ad0d7..34776493 100644
--- a/src/requests/utils.py
+++ b/src/requests/utils.py
@@ -851,7 +851,7 @@ def should_bypass_proxies(url: str, no_proxy: str | None) -> bool:
                 host_with_port += f":{parsed.port}"
 
             for host in no_proxy_hosts:
-                if hostname.endswith(host) or host_with_port.endswith(host):
+                if hostname == host or hostname.endswith("." + host) or host_with_port == host or host_with_port.endswith("." + host):
                     # The URL does match something in no_proxy, so we don't want
                     # to apply the proxies on this URL.
                     return True

Test output

show
..s..................................................................... [ 31%]
........................................................................ [ 62%]
....................................................................ssss [ 94%]
sssssss.....s                                                            [100%]
=============================== warnings summary ===============================
../../../../../../kaggle/tmp/envs/requests/lib/python3.13/site-packages/_pytest/config/__init__.py:1464
  /kaggle/tmp/envs/requests/lib/python3.13/site-packages/_pytest/config/__init__.py:1464: PytestConfigWarning: Unknown config option: timeout
  
    self._warn_or_fail_if_strict(f"Unknown config option: {key}\n")

tests/test_utils.py::TestContentEncodingDetection::test_none
tests/test_utils.py::TestContentEncodingDetection::test_pragmas[<meta charset="UTF-8">]
tests/test_utils.py::TestContentEncodingDetection::test_pragmas[<meta http-equiv="Content-type" content="text/html;charset=UTF-8">]
tests/test_utils.py::TestContentEncodingDetection::test_pragmas[<meta http-equiv="Content-type" content="text/html;charset=UTF-8" />]
tests/test_utils.py::TestContentEncodingDetection::test_pragmas[<?xml version="1.0" encoding="UTF-8"?>]
tests/test_utils.py::TestContentEncodingDetection::test_precedence
  /tmp/swe_work/eval3_submission_v2/requests_7427/b/workspace/src/requests/utils.py:527: DeprecationWarning: In requests 3.0, get_encodings_from_content will be removed. For more information, please see the discussion on issue #2266. (This warning should only appear once.)
    warnings.warn(

-- Docs: https://docs.pytest.org/en/stable/how-to/capture-warnings.html
216 passed, 13 skipped, 7 warnings in 0.24s
[2026-09-24 10:48:11,245] WARNING in core: flasgger is not installed; serving the static landing page at / and skipping the Swagger UI and /spec.json.