Bug #21790: `Socket.getaddrinfo` hangs after `fork()` on macOS 26.1 (Tahoe) for IPv4-only hosts - Ruby - Ruby Issue Tracking System

Actions

Copy link

Bug #21790

closed

`Socket.getaddrinfo` hangs after `fork()` on macOS 26.1 (Tahoe) for IPv4-only hosts

Bug #21790: `Socket.getaddrinfo` hangs after `fork()` on macOS 26.1 (Tahoe) for IPv4-only hosts

Added by adamoffat (Adam Moffat) 3 months ago. Updated 2 months ago.

Status:

Third Party's Issue

Assignee:

Target version:

ruby -v:

3.3.8

Backport:

3.2: UNKNOWN, 3.3: UNKNOWN, 3.4: UNKNOWN

[ruby-core:124288]

Description

Ruby's Socket.getaddrinfo hangs indefinitely in forked child processes on macOS 26.1 (Tahoe) when resolving IPv4-only hostnames. This is a regression that does not occur on macOS 15.x (Sonoma) or earlier.

Ruby version:
ruby 3.3.8 (2025-04-09 revision b200bad6cd) [arm64-darwin24]
Also confirmed this affects Ruby 3.2.6 and 3.4.1.

Reproducible script:

require "socket"
require "timeout"

puts "Ruby #{RUBY_VERSION} on #{RUBY_PLATFORM}"
Socket.getaddrinfo("api.segment.io", 443, nil, :STREAM)
puts "Parent: DNS completed"

pid = fork do
  puts "Child: Attempting DNS resolution..."
  begin
    Timeout.timeout(90) do
      Socket.getaddrinfo("api.segment.io", 443, nil, :STREAM)
    end
    puts "Child: SUCCESS"
    exit 0
  rescue Timeout::Error
      puts "Child: FAILED - hung for 90 seconds"
      exit 1
  end
end

Process.wait(pid)

Note: Remove the Timeout.timeout(90) wrapper to observe the hang indefinitely. The timeout is included only to allow the script to exit for testing purposes.

Result of reproduce process:

Ruby 3.3.8 on arm64-darwin24
Parent: DNS completed
Child: Attempting DNS resolution...
Child: FAILED - hung for 90 seconds

The child process hangs with one thread consuming 100% CPU.

Expected result: The child process should complete DNS resolution successfully, as it does on macOS 15.x and earlier.

Analysis:
Stack trace shows:
Main thread: Blocked in wait_getaddrinfo → _pthread_cond_wait
DNS thread: Spinning in _gai_nat64_second_pass → nw_path_access_agent_cache → _os_log_preferences_refresh → SIGSEGV

The crash occurs in macOS's NAT64 synthesis code path. Ruby's signal handler catches the SIGSEGV but cannot recover, causing the DNS thread to spin.

Key observations:

Only affects IPv4-only hosts. Hosts with IPv6 (like google.com) work correctly.
Using AF_INET instead of AF_UNSPEC works. Socket.getaddrinfo("api.segment.io", 443, Socket::AF_INET, :STREAM) succeeds.
Python is not affected. Python calls getaddrinfo() synchronously without a background thread.
Parent must do DNS before fork. If the parent has not called getaddrinfo(), the child works correctly.

Workaround:

Use resolv-replace to bypass the native DNS resolver: require "resolv-replace"

Impact:
This breaks all Ruby applications using pre-forking worker models (Resque, Unicorn, Puma, Sidekiq, Passenger) on macOS Tahoe.

Apple Bug Report:
Filed with Apple as Feedback Assistant #FB21364061

Files

Download all files

stack_trace.txt (66.6 KB) stack_trace.txt	Stack Trace for Bug	adamoffat (Adam Moffat), 12/17/2025 05:56 PM
ruby_dns_fork_bug.rb (1.02 KB) ruby_dns_fork_bug.rb	Reproduction Script	adamoffat (Adam Moffat), 12/17/2025 06:02 PM
ruby_3.2.6_crash_output.txt (1.79 KB) ruby_3.2.6_crash_output.txt	Ruby 3.2.6 stacktrace	adamoffat (Adam Moffat), 12/18/2025 03:44 PM
python_dns_fork_test.py (1.8 KB) python_dns_fork_test.py	Python reproduction script	adamoffat (Adam Moffat), 12/18/2025 06:26 PM
python_dns_fork_test.py (1.97 KB) python_dns_fork_test.py	Python Reproduction Script	adamoffat (Adam Moffat), 12/18/2025 06:39 PM
python_crash_output.txt (1.28 KB) python_crash_output.txt	Python Crash output	adamoffat (Adam Moffat), 12/18/2025 06:39 PM

Related issues 2 (0 open — 2 closed)

Updated by adamoffat (Adam Moffat) 3 months ago Actions
Copy link
#1 [ruby-core:124296]

To confirm: MacOS Sequoia also did not have this issue.

Updated by adamoffat (Adam Moffat) 3 months ago Actions
Copy link
#2 [ruby-core:124297]

I saw that this was added in 3.4.0: https://github.com/ruby/ruby/pull/10864

Seen here: (https://github.com/ruby/ruby/releases/tag/v3_4_0_preview2)

But I also tested this using 3.4.1 and it was still an issue.

Updated by mame (Yusuke Endoh) 3 months ago Actions
Copy link
#3 [ruby-core:124299]

Thank you for the report.

Since I don't have access to Tahoe, I cannot test this in my own environment. However, I have a few questions to clarify the situation.

The change to perform DNS lookups in a dedicated background thread was introduced in Ruby 3.3.0. You mentioned that this affects Ruby 3.2.6 as well. Are you certain it reproduces on 3.2.6?

If it fails on 3.2.6, the cause might be unrelated to the background thread, as its behavior should be similar to Python's. Would it be possible to provide a stack trace from the 3.2.6 crash?

Though it's just a guess, this might be a bug with getaddrinfo on Tahoe itself, but I could be wrong.

Updated by adamoffat (Adam Moffat) 3 months ago · Edited Actions
Copy link
#4 [ruby-core:124303]

File ruby_3.2.6_crash_output.txt ruby_3.2.6_crash_output.txt added

mame (Yusuke Endoh) wrote in #note-3:

Thank you for the report.

Since I don't have access to Tahoe, I cannot test this in my own environment. However, I have a few questions to clarify the situation.

The change to perform DNS lookups in a dedicated background thread was introduced in Ruby 3.3.0. You mentioned that this affects Ruby 3.2.6 as well. Are you certain it reproduces on 3.2.6?

If it fails on 3.2.6, the cause might be unrelated to the background thread, as its behavior should be similar to Python's. Would it be possible to provide a stack trace from the 3.2.6 crash?

Though it's just a guess, this might be a bug with getaddrinfo on Tahoe itself, but I could be wrong.

Ah yes, sorry I should have clarified this in my post. I tested this in 3.2.6 but it manifests differently in that version.

When I ran the same reproduction script with Ruby 3.2.6, rather than hanging indefinitely, it crashed immediately with a segmentation fault when the child process attempts DNS resolution.

The crash occurs at the getaddrinfo call in the forked child. The backtrace shows the fault originating in macOS system libraries, specifically in libsystem_trace.dylib at _os_log_preferences_refresh.

This confirms Ruby 3.2.6 is also affected by the same underlying issue - it just manifests as an immediate crash rather than a hang.

I've attached the full crash output for reference.

Updated by mame (Yusuke Endoh) 3 months ago Actions
Copy link
#5 [ruby-core:124306]

Thank you. This looks like the same issue reported multiple times in the past, but we were previously stuck without a way to investigate.

https://bugs.ruby-lang.org/issues/15490
https://bugs.ruby-lang.org/issues/15794
https://github.com/redis/redis-rb/issues/859
https://github.com/hanami/hanami/issues/993

It is greatly appreciated that the reproduction conditions are now much clearer.

This issue does not affect Python even in a forked child process, right? If Python avoids this error, checking how it calls getaddrinfo might give us a hint for a fix or workaround.

It is difficult for me to debug this without a reproducing environment. Are there any committers or contributors who can reproduce the issue and investigate?

Updated by mame (Yusuke Endoh) 3 months ago Actions
Copy link
#6

Related to Bug #15490: socket.rb - recurring segmentation faults added
Related to Bug #15794: Can not start Puma with Rails after bundle install added

Updated by adamoffat (Adam Moffat) 3 months ago · Edited Actions
Copy link
#7 [ruby-core:124308]

File python_dns_fork_test.py python_dns_fork_test.py added

Updated by adamoffat (Adam Moffat) 3 months ago Actions
Copy link Download all files
#8 [ruby-core:124309]

File python_dns_fork_test.py python_dns_fork_test.py added
File python_crash_output.txt python_crash_output.txt added

Ah my earlier Python script had a bug.

My initial Python test incorrectly reported success. The script used os.WEXITSTATUS() to check the child's exit status, but this function only works for processes that exit normally. When a process is killed by a signal (SIGSEGV), it returns 0, giving a false positive.

After fixing the script to check os.WIFSIGNALED(), I was able to confirm the child is killed by signal 11 (SIGSEGV). The crash logs show the identical stack trace to Ruby: _gai_nat64_second_pass → nw_path_access_agent_cache → _os_log_preferences_refresh.

This is an OS-level bug in macOS Tahoe, not language-specific. My apologies.

Updated by mame (Yusuke Endoh) 2 months ago Actions
Copy link
#9 [ruby-core:124514]

Status changed from Open to Third Party's Issue

Thank you for your confirmation. This is most likely a macOS bug, so I'd close this as a third-party issue.
It would be the best for macOS to fix the issue, but if someone finds a workaround, I'd consider importing it in the Ruby side.

Actions

Copy link

Also available in: PDF Atom

Project

General

Profile

Ruby

Custom queries

Bug #21790

`Socket.getaddrinfo` hangs after `fork()` on macOS 26.1 (Tahoe) for IPv4-only hosts

Updated by adamoffat (Adam Moffat) 3 months ago Actions
Copy link
#1 [ruby-core:124296]

Updated by adamoffat (Adam Moffat) 3 months ago Actions
Copy link
#2 [ruby-core:124297]

Updated by mame (Yusuke Endoh) 3 months ago Actions
Copy link
#3 [ruby-core:124299]

Updated by adamoffat (Adam Moffat) 3 months ago · Edited Actions
Copy link
#4 [ruby-core:124303]

Updated by mame (Yusuke Endoh) 3 months ago Actions
Copy link
#5 [ruby-core:124306]

Updated by mame (Yusuke Endoh) 3 months ago Actions
Copy link
#6

Updated by adamoffat (Adam Moffat) 3 months ago · Edited Actions
Copy link
#7 [ruby-core:124308]

Updated by adamoffat (Adam Moffat) 3 months ago Actions
Copy link Download all files
#8 [ruby-core:124309]

Updated by mame (Yusuke Endoh) 2 months ago Actions
Copy link
#9 [ruby-core:124514]

Project

General

Profile

Ruby

Custom queries

Bug #21790

`Socket.getaddrinfo` hangs after `fork()` on macOS 26.1 (Tahoe) for IPv4-only hosts

Updated by adamoffat (Adam Moffat) 3 months ago ActionsCopy link #1 [ruby-core:124296]

Updated by adamoffat (Adam Moffat) 3 months ago ActionsCopy link #2 [ruby-core:124297]

Updated by mame (Yusuke Endoh) 3 months ago ActionsCopy link #3 [ruby-core:124299]

Updated by adamoffat (Adam Moffat) 3 months ago · Edited ActionsCopy link #4 [ruby-core:124303]

Updated by mame (Yusuke Endoh) 3 months ago ActionsCopy link #5 [ruby-core:124306]

Updated by mame (Yusuke Endoh) 3 months ago ActionsCopy link #6

Updated by adamoffat (Adam Moffat) 3 months ago · Edited ActionsCopy link #7 [ruby-core:124308]

Updated by adamoffat (Adam Moffat) 3 months ago ActionsCopy link Download all files #8 [ruby-core:124309]

Updated by mame (Yusuke Endoh) 2 months ago ActionsCopy link #9 [ruby-core:124514]

Updated by adamoffat (Adam Moffat) 3 months ago Actions
Copy link
#1 [ruby-core:124296]

Updated by adamoffat (Adam Moffat) 3 months ago Actions
Copy link
#2 [ruby-core:124297]

Updated by mame (Yusuke Endoh) 3 months ago Actions
Copy link
#3 [ruby-core:124299]

Updated by adamoffat (Adam Moffat) 3 months ago · Edited Actions
Copy link
#4 [ruby-core:124303]

Updated by mame (Yusuke Endoh) 3 months ago Actions
Copy link
#5 [ruby-core:124306]

Updated by mame (Yusuke Endoh) 3 months ago Actions
Copy link
#6

Updated by adamoffat (Adam Moffat) 3 months ago · Edited Actions
Copy link
#7 [ruby-core:124308]

Updated by adamoffat (Adam Moffat) 3 months ago Actions
Copy link Download all files
#8 [ruby-core:124309]

Updated by mame (Yusuke Endoh) 2 months ago Actions
Copy link
#9 [ruby-core:124514]