We are currently developing a Ruby based web application which connects to a DB2 Database and we have been using ibm_db-5.4.0 to establish a connection, suddenly we got a error related to RUBY garbage collector PHASE.
We have checked the issue with IBM_team to make sure that It was not a IBM_GEM problem but as a result of their tests, IBM_GEM is working in different cases but for us we face up with those errors even with those versions (2.7.6, 3.1.2, 3.2.1):
Isn't this a duplicate of #19524? I don't think you will get a different answer to this ticket from the one that was given in that ticket. The bug is indicating that a C-extension (probably, the IBM DB Adapter) is allocating during garbage collection phase.
Unfortunately, the traces are not very helpful.
But I agree, it's likely an issue with the IBM gem
and not a Ruby issue.
Here is a guess as to what's happening: the traces indicate that it's crashing soon after a new Ruby thread starts up,
and for "Object allocation during garbage collection phase" to happen at that timing,
maybe someone is running Ruby code without holding the global VM lock. Any code paths in
the gem that terminates in rb_thread_call_without_gvl(), such as ones involving _ruby_ibm_db_check_sql_errors(), can have this class of bugs.
Hi everyone.
As I mentioned the first thing I did was check the issue with IBM, if you check the stack message, our application (PEC) is started, then it establish the connection to the database using IBM_GEM and this error is triggered by ruby after the ibm gem connects to the database.
[INFO] Process pid 1002200
[INFO] Initializing PEC Engine ...
[INFO] Establishing connection to control database ...
[INFO] Engine started
/ruby-3.2.1/lib/ruby/3.2.0/rubygems/specification.rb:1048: [BUG] object allocation during garbage collection phase ruby 3.2.1.
Our application is something like a process dispatcher for a few moments it runs very well but after that suddenly ruby crashes.
about your guess that someone else is running Ruby code without holding the global VM lock, is not possible because in our Centos 8 server is the only application running and we have enough physical resources to run.
Please understand that while Ruby giving you the crash report
saying [BUG] object allocation during garbage collection phase
is the proximate cause of the problem, I'm telling you that
the ultimate cause is likely a bug in the IBM gem. C extensions
have contracts they must uphold, and if they don't, they
crash the whole process. Ruby is the messenger in all crashes,
but it is incorrect to always blame the messenger. Crashes
involving threads like the one you are facing are by nature non-local;
the code initiating the crash is not always problematic.
about your guess that someone else is running Ruby code without holding the global VM lock, is not possible because in our Centos 8 server is the only application running and we have enough physical resources to run.
It is absolutely possible. The VM lock arbitrates
concurrency of threads within the same process. From the
stack trace, it's clear that you have multiple threads in your
Ruby process. The lock has nothing to do with how many
applications you run on your server. If you have multiple threads
in your Ruby process, it's in-play.
Repeat (2) to rebuild and reinstall the gem now that it's changed
See if the crash still reproduces.
If this patch makes the crash go away, we can say with high confidence
that the ibm_db gem is misusing rb_thread_call_without_gvl(). Send this
to IBM as a bug report.
If the crash still happens, maybe you can try reproducing the bug without
any third-party C extensions. If you can do that, that'd be a more
actionable bug report for us. There is not much we can do on our end
with the information you have posted.