Groups | Search | Server Info | Keyboard shortcuts | Login | Register [http] [https] [nntp] [nntps]


Groups > linux.kernel > #1711713 > unrolled thread

Re: [PATCH] f2fs: let fill_super handle roll-forward errors

Started byChao Yu <yuchao0@huawei.com>
First post2017-08-15 03:50 +0200
Last post2017-08-16 03:20 +0200
Articles 5 — 2 participants

Back to article view | Back to linux.kernel

This discussion starts older than the indexed window; earlier articles aren't shown. The article labeled Started by below is the oldest one visible, not the original post.


Contents

  Re: [PATCH] f2fs: let fill_super handle roll-forward errors Chao Yu <yuchao0@huawei.com> - 2017-08-15 03:50 +0200
    Re: [PATCH] f2fs: let fill_super handle roll-forward errors Jaegeuk Kim <jaegeuk@kernel.org> - 2017-08-15 05:30 +0200
      Re: [PATCH] f2fs: let fill_super handle roll-forward errors Chao Yu <yuchao0@huawei.com> - 2017-08-15 13:50 +0200
        Re: [PATCH] f2fs: let fill_super handle roll-forward errors Jaegeuk Kim <jaegeuk@kernel.org> - 2017-08-15 18:50 +0200
          Re: [PATCH] f2fs: let fill_super handle roll-forward errors Chao Yu <yuchao0@huawei.com> - 2017-08-16 03:20 +0200

#1711713 — Re: [PATCH] f2fs: let fill_super handle roll-forward errors

FromChao Yu <yuchao0@huawei.com>
Date2017-08-15 03:50 +0200
SubjectRe: [PATCH] f2fs: let fill_super handle roll-forward errors
Message-ID<uez9x-3gj-53@gated-at.bofh.it>
Hi Jaegeuk,

On 2017/8/11 8:42, Jaegeuk Kim wrote:
> If we set CP_ERROR_FLAG in roll-forward error, f2fs is no longer to proceed
> any IOs due to f2fs_cp_error(). But, for example, if some stale data is involved
> on roll-forward process, we're able to get -ENOENT, getting fs stuck.
> If we get any error, let fill_super set SBI_NEED_FSCK and try to recover back
> to stable point.

Before that, we have cleaned up all node/meta page cache, so we will get back to
last checkpoint status, means losing fsynced datas for ever.

Would it be better to just leave message reminding user to mount with
disable_roll_forward or run fsck offline.

Thanks,

> 
> Cc: <stable@vger.kernel.org>
> Signed-off-by: Jaegeuk Kim <jaegeuk@kernel.org>
> ---
>  fs/f2fs/recovery.c | 2 --
>  1 file changed, 2 deletions(-)
> 
> diff --git a/fs/f2fs/recovery.c b/fs/f2fs/recovery.c
> index a3d02613934a..f707d810c87d 100644
> --- a/fs/f2fs/recovery.c
> +++ b/fs/f2fs/recovery.c
> @@ -649,8 +649,6 @@ int recover_fsync_data(struct f2fs_sb_info *sbi, bool check_only)
>  	}
>  
>  	clear_sbi_flag(sbi, SBI_POR_DOING);
> -	if (err)
> -		set_ckpt_flags(sbi, CP_ERROR_FLAG);
>  	mutex_unlock(&sbi->cp_mutex);
>  
>  	/* let's drop all the directory inodes for clean checkpoint */
> 

[toc] | [next] | [standalone]


#1711809

FromJaegeuk Kim <jaegeuk@kernel.org>
Date2017-08-15 05:30 +0200
Message-ID<ueAIi-4oo-11@gated-at.bofh.it>
In reply to#1711713
On 08/15, Chao Yu wrote:
> Hi Jaegeuk,
> 
> On 2017/8/11 8:42, Jaegeuk Kim wrote:
> > If we set CP_ERROR_FLAG in roll-forward error, f2fs is no longer to proceed
> > any IOs due to f2fs_cp_error(). But, for example, if some stale data is involved
> > on roll-forward process, we're able to get -ENOENT, getting fs stuck.
> > If we get any error, let fill_super set SBI_NEED_FSCK and try to recover back
> > to stable point.
> 
> Before that, we have cleaned up all node/meta page cache, so we will get back to
> last checkpoint status, means losing fsynced datas for ever.
> 
> Would it be better to just leave message reminding user to mount with
> disable_roll_forward or run fsck offline.

We can't rely on user for this, since fsck cannot recover this, resulting in
infinite mount failure. The only way is to disable roll-forward recovery, which
is same as returning error here.

Thanks,

> 
> Thanks,
> 
> > 
> > Cc: <stable@vger.kernel.org>
> > Signed-off-by: Jaegeuk Kim <jaegeuk@kernel.org>
> > ---
> >  fs/f2fs/recovery.c | 2 --
> >  1 file changed, 2 deletions(-)
> > 
> > diff --git a/fs/f2fs/recovery.c b/fs/f2fs/recovery.c
> > index a3d02613934a..f707d810c87d 100644
> > --- a/fs/f2fs/recovery.c
> > +++ b/fs/f2fs/recovery.c
> > @@ -649,8 +649,6 @@ int recover_fsync_data(struct f2fs_sb_info *sbi, bool check_only)
> >  	}
> >  
> >  	clear_sbi_flag(sbi, SBI_POR_DOING);
> > -	if (err)
> > -		set_ckpt_flags(sbi, CP_ERROR_FLAG);
> >  	mutex_unlock(&sbi->cp_mutex);
> >  
> >  	/* let's drop all the directory inodes for clean checkpoint */
> > 

[toc] | [prev] | [next] | [standalone]


#1712109

FromChao Yu <yuchao0@huawei.com>
Date2017-08-15 13:50 +0200
Message-ID<ueIwa-Mq-9@gated-at.bofh.it>
In reply to#1711809
On 2017/8/15 11:22, Jaegeuk Kim wrote:
> On 08/15, Chao Yu wrote:
>> Hi Jaegeuk,
>>
>> On 2017/8/11 8:42, Jaegeuk Kim wrote:
>>> If we set CP_ERROR_FLAG in roll-forward error, f2fs is no longer to proceed
>>> any IOs due to f2fs_cp_error(). But, for example, if some stale data is involved
>>> on roll-forward process, we're able to get -ENOENT, getting fs stuck.
>>> If we get any error, let fill_super set SBI_NEED_FSCK and try to recover back
>>> to stable point.
>>
>> Before that, we have cleaned up all node/meta page cache, so we will get back to
>> last checkpoint status, means losing fsynced datas for ever.
>>
>> Would it be better to just leave message reminding user to mount with
>> disable_roll_forward or run fsck offline.
> 
> We can't rely on user for this, since fsck cannot recover this, resulting in

If fsck has no ability to recover this, it could tag superblock in somewhere,
then kernel could skip recovery. Comparing to fail recovery directly, it give
user another chance to rescuer his datas.

Thanks,

> infinite mount failure. The only way is to disable roll-forward recovery, which
> is same as returning error here.
> 
> Thanks,
> 
>>
>> Thanks,
>>
>>>
>>> Cc: <stable@vger.kernel.org>
>>> Signed-off-by: Jaegeuk Kim <jaegeuk@kernel.org>
>>> ---
>>>  fs/f2fs/recovery.c | 2 --
>>>  1 file changed, 2 deletions(-)
>>>
>>> diff --git a/fs/f2fs/recovery.c b/fs/f2fs/recovery.c
>>> index a3d02613934a..f707d810c87d 100644
>>> --- a/fs/f2fs/recovery.c
>>> +++ b/fs/f2fs/recovery.c
>>> @@ -649,8 +649,6 @@ int recover_fsync_data(struct f2fs_sb_info *sbi, bool check_only)
>>>  	}
>>>  
>>>  	clear_sbi_flag(sbi, SBI_POR_DOING);
>>> -	if (err)
>>> -		set_ckpt_flags(sbi, CP_ERROR_FLAG);
>>>  	mutex_unlock(&sbi->cp_mutex);
>>>  
>>>  	/* let's drop all the directory inodes for clean checkpoint */
>>>
> 
> .
> 

[toc] | [prev] | [next] | [standalone]


#1712342

FromJaegeuk Kim <jaegeuk@kernel.org>
Date2017-08-15 18:50 +0200
Message-ID<ueNcu-3Ii-13@gated-at.bofh.it>
In reply to#1712109
On 08/15, Chao Yu wrote:
> On 2017/8/15 11:22, Jaegeuk Kim wrote:
> > On 08/15, Chao Yu wrote:
> >> Hi Jaegeuk,
> >>
> >> On 2017/8/11 8:42, Jaegeuk Kim wrote:
> >>> If we set CP_ERROR_FLAG in roll-forward error, f2fs is no longer to proceed
> >>> any IOs due to f2fs_cp_error(). But, for example, if some stale data is involved
> >>> on roll-forward process, we're able to get -ENOENT, getting fs stuck.
> >>> If we get any error, let fill_super set SBI_NEED_FSCK and try to recover back
> >>> to stable point.
> >>
> >> Before that, we have cleaned up all node/meta page cache, so we will get back to
> >> last checkpoint status, means losing fsynced datas for ever.
> >>
> >> Would it be better to just leave message reminding user to mount with
> >> disable_roll_forward or run fsck offline.
> > 
> > We can't rely on user for this, since fsck cannot recover this, resulting in
> 
> If fsck has no ability to recover this, it could tag superblock in somewhere,
> then kernel could skip recovery. Comparing to fail recovery directly, it give
> user another chance to rescuer his datas.

Huh, what do you mean? This patch let f2fs_fill_super set SBI_NEED_FSCK in
superblock and skip roll-forward in second round.

> 
> Thanks,
> 
> > infinite mount failure. The only way is to disable roll-forward recovery, which
> > is same as returning error here.
> > 
> > Thanks,
> > 
> >>
> >> Thanks,
> >>
> >>>
> >>> Cc: <stable@vger.kernel.org>
> >>> Signed-off-by: Jaegeuk Kim <jaegeuk@kernel.org>
> >>> ---
> >>>  fs/f2fs/recovery.c | 2 --
> >>>  1 file changed, 2 deletions(-)
> >>>
> >>> diff --git a/fs/f2fs/recovery.c b/fs/f2fs/recovery.c
> >>> index a3d02613934a..f707d810c87d 100644
> >>> --- a/fs/f2fs/recovery.c
> >>> +++ b/fs/f2fs/recovery.c
> >>> @@ -649,8 +649,6 @@ int recover_fsync_data(struct f2fs_sb_info *sbi, bool check_only)
> >>>  	}
> >>>  
> >>>  	clear_sbi_flag(sbi, SBI_POR_DOING);
> >>> -	if (err)
> >>> -		set_ckpt_flags(sbi, CP_ERROR_FLAG);
> >>>  	mutex_unlock(&sbi->cp_mutex);
> >>>  
> >>>  	/* let's drop all the directory inodes for clean checkpoint */
> >>>
> > 
> > .
> > 

[toc] | [prev] | [next] | [standalone]


#1712567

FromChao Yu <yuchao0@huawei.com>
Date2017-08-16 03:20 +0200
Message-ID<ueVa1-qT-3@gated-at.bofh.it>
In reply to#1712342
On 2017/8/16 0:42, Jaegeuk Kim wrote:
> On 08/15, Chao Yu wrote:
>> On 2017/8/15 11:22, Jaegeuk Kim wrote:
>>> On 08/15, Chao Yu wrote:
>>>> Hi Jaegeuk,
>>>>
>>>> On 2017/8/11 8:42, Jaegeuk Kim wrote:
>>>>> If we set CP_ERROR_FLAG in roll-forward error, f2fs is no longer to proceed
>>>>> any IOs due to f2fs_cp_error(). But, for example, if some stale data is involved
>>>>> on roll-forward process, we're able to get -ENOENT, getting fs stuck.
>>>>> If we get any error, let fill_super set SBI_NEED_FSCK and try to recover back
>>>>> to stable point.
>>>>
>>>> Before that, we have cleaned up all node/meta page cache, so we will get back to
>>>> last checkpoint status, means losing fsynced datas for ever.
>>>>
>>>> Would it be better to just leave message reminding user to mount with
>>>> disable_roll_forward or run fsck offline.
>>>
>>> We can't rely on user for this, since fsck cannot recover this, resulting in
>>
>> If fsck has no ability to recover this, it could tag superblock in somewhere,
>> then kernel could skip recovery. Comparing to fail recovery directly, it give
>> user another chance to rescuer his datas.
> 
> Huh, what do you mean? This patch let f2fs_fill_super set SBI_NEED_FSCK in
> superblock and skip roll-forward in second round.

Oh, just notice that additional checkpoint which sets SBI_NEED_FSCK won't
destroy warn node chain, so we have chance to run fsck and recover after another
mount. Sorry.

Reviewed-by: Chao Yu <yuchao0@huawei.com>

> 
>>
>> Thanks,
>>
>>> infinite mount failure. The only way is to disable roll-forward recovery, which
>>> is same as returning error here.
>>>
>>> Thanks,
>>>
>>>>
>>>> Thanks,
>>>>
>>>>>
>>>>> Cc: <stable@vger.kernel.org>
>>>>> Signed-off-by: Jaegeuk Kim <jaegeuk@kernel.org>
>>>>> ---
>>>>>  fs/f2fs/recovery.c | 2 --
>>>>>  1 file changed, 2 deletions(-)
>>>>>
>>>>> diff --git a/fs/f2fs/recovery.c b/fs/f2fs/recovery.c
>>>>> index a3d02613934a..f707d810c87d 100644
>>>>> --- a/fs/f2fs/recovery.c
>>>>> +++ b/fs/f2fs/recovery.c
>>>>> @@ -649,8 +649,6 @@ int recover_fsync_data(struct f2fs_sb_info *sbi, bool check_only)
>>>>>  	}
>>>>>  
>>>>>  	clear_sbi_flag(sbi, SBI_POR_DOING);
>>>>> -	if (err)
>>>>> -		set_ckpt_flags(sbi, CP_ERROR_FLAG);
>>>>>  	mutex_unlock(&sbi->cp_mutex);
>>>>>  
>>>>>  	/* let's drop all the directory inodes for clean checkpoint */
>>>>>
>>>
>>> .
>>>
> 
> .
> 

[toc] | [prev] | [standalone]


Back to top | Article view | linux.kernel


csiph-web