Intel製NICドライバ(e1000e)のパケット詰まり
リモートで操作していたサーバーが作業終了したのでexitした後に気になってアクセスしたら繋がらない・・・夜中だったので諦めてその日は寝たのだけど、朝になってもやはり回復しない。
で・・・週末まで待ったやっとサーバーの置いてある実家に来たのだけど
[Sat Aug 15 15:38:29 2026] e1000e 0000:00:19.0 eno1: Detected Hardware Unit Hang:
TDH <8>
TDT <83>
next_to_use <83>
next_to_clean <8>
buffer_info[next_to_clean]:
time_stamp <100205561>
next_to_watch <8>
jiffies <100207781>
next_to_watch.status <0>
MAC Status <40080083>
PHY Status <796d>
PHY 1000BASE-T Status <3800>
PHY Extended Status <3000>
PCI Status <10>
[Sat Aug 15 15:38:31 2026] Call Trace:
[Sat Aug 15 15:38:31 2026] <IRQ>
[Sat Aug 15 15:38:31 2026] ? __warn+0x94/0xe0
[Sat Aug 15 15:38:31 2026] ? dev_watchdog+0x29a/0x2b0
[Sat Aug 15 15:38:31 2026] ? dev_watchdog+0x29a/0x2b0
[Sat Aug 15 15:38:31 2026] ? report_bug+0xb1/0xe0
[Sat Aug 15 15:38:31 2026] ? do_error_trap+0x9e/0xd0
[Sat Aug 15 15:38:31 2026] ? do_invalid_op+0x36/0x40
[Sat Aug 15 15:38:31 2026] ? dev_watchdog+0x29a/0x2b0
[Sat Aug 15 15:38:31 2026] ? invalid_op+0x14/0x20
[Sat Aug 15 15:38:31 2026] ? dev_watchdog+0x29a/0x2b0
[Sat Aug 15 15:38:31 2026] ? dev_watchdog+0x29a/0x2b0
[Sat Aug 15 15:38:31 2026] ? pfifo_fast_enqueue+0x150/0x150
[Sat Aug 15 15:38:31 2026] call_timer_fn+0x2e/0x130
[Sat Aug 15 15:38:31 2026] run_timer_softirq+0x1e5/0x440
[Sat Aug 15 15:38:31 2026] ? sched_clock+0x5/0x10
[Sat Aug 15 15:38:31 2026] __do_softirq+0xdc/0x2cf
[Sat Aug 15 15:38:31 2026] irq_exit_rcu+0xc6/0xd0
[Sat Aug 15 15:38:31 2026] irq_exit+0xa/0x10
[Sat Aug 15 15:38:31 2026] smp_apic_timer_interrupt+0x74/0x130
[Sat Aug 15 15:38:31 2026] apic_timer_interrupt+0xf/0x20
[Sat Aug 15 15:38:31 2026] </IRQ>
[Sat Aug 15 15:38:31 2026] RIP: 0033:0x7fe5e49732e5
[Sat Aug 15 15:38:31 2026] Code: c1 81 e1 7f 01 00 00 83 e0 7c 75 49 41 89 c0 85 c9 75 4a c7 42 08 00 00 00 00 85 f6 75 2f 8b 72 10 81 e6 80 00 00 00 f0 ff 0a <74> 16 48 8d 3a 48 81 ec 80 00 00 00 e8 da 55 00 00 48 81 c4 80 00
[Sat Aug 15 15:38:31 2026] RSP: 002b:00007fffd1bbb318 EFLAGS: 00000246 ORIG_RAX: ffffffffffffff13
[Sat Aug 15 15:38:31 2026] RAX: 0000000000000000 RBX: 000055cc193c4f60 RCX: 0000000000000000
[Sat Aug 15 15:38:31 2026] RDX: 000055cc19217b78 RSI: 0000000000000000 RDI: 000055cc19217b78
[Sat Aug 15 15:38:31 2026] RBP: 000055cc19217b78 R08: 0000000000000000 R09: 000000000048a008
[Sat Aug 15 15:38:31 2026] R10: 000000000048a008 R11: 0000000000000246 R12: 0000000000000000
[Sat Aug 15 15:38:31 2026] R13: 00000225ff07bd33 R14: 000055cc175f50f0 R15: 000055cc19217b70
[Sat Aug 15 15:38:31 2026] ---[ end trace ac47890899668efd ]---
[Sat Aug 15 15:38:31 2026] e1000e 0000:00:19.0 eno1: Reset adapter unexpectedly
って感じのエラーでネットワークがハングしてる。 Geminiさんに聞いてみたところ、Intel製NICドライバ(e1000e)の典型的なパケット詰まり(TX Timeout)とNICのリセットが発生しているそうな
エラーの状況と原因
dev_watchdog / pfifo_fast_enqueue: カーネルのネットワーク監視機能が「パケット送信のタイムアウト(ハングアップ)」を検知しました。
e1000e … Reset adapter unexpectedly: ドライバがNICを復旧させるため、eno1 を強制的に再起動しました。
br0: port 1(eno1) entered disabled/forwarding state: NICリセットに伴い、ブリッジ(br0)のポートが一瞬切断された後、再接続されています。
なぜ発生するのか?
IntelのオンボードLAN(I217 / I218 / I219 シリーズなど)とLinuxの e1000e ドライバの組み合わせにおいて、ハードウェアオフロード機能(TSO) や 省電力機能(EEE / ASPM) の互換性問題により、負荷がかかった際や特定の通信時にパケットが詰まる公知の不具合です。
対応方法
一時的な対応
sudo ethtool -K eno1 tso off gso off gro off
sudo ethtool --set-eee eno1 eee off
恒久的な対応
sudo nmcli connection modify eno1 ethtool.feature-tso off
sudo nmcli connection modify eno1 ethtool.feature-gso off
sudo nmcli connection modify eno1 ethtool.feature-gro off
sudo nmcli connection up eno1
設定が反映されたか確認する
ethtool -k eno1 | grep -E "tcp-segmentation-offload|generic-segmentation-offload|generic-receive-offload"