如何在文件夹层次结构中找到所有不同的文件扩展名?

在Linux机器上，我希望遍历一个文件夹层次结构，并获得其中所有不同文件扩展名的列表。

从外壳中实现这一点的最佳方法是什么?

当前回答

到目前为止，没有一个回复正确地处理带有换行符的文件名(除了ChristopheD的，它是在我输入这篇文章时出现的)。下面不是一个shell一行程序，但是可以工作，并且相当快。

import os, sys

def names(roots):
    for root in roots:
        for a, b, basenames in os.walk(root):
            for basename in basenames:
                yield basename

sufs = set(os.path.splitext(x)[1] for x in names(sys.argv[1:]))
for suf in sufs:
    if suf:
        print suf

2009-12-04 08:35:28

其他回答

试试这个(不确定是不是最好的方法，但确实有效):

find . -type f | perl -ne 'print $1 if m/\.([^.\/]+)$/' | sort -u

它的工作原理如下:

找到当前文件夹中的所有文件打印文件扩展名(如果有的话) 制作一个唯一的排序列表

2009-12-03 19:21:46

不需要管道排序，awk可以做所有的事情:

find . -type f | awk -F. '!a[$NF]++{print $NF}'

2011-08-24 05:21:12

递归版本:

find . -type f | sed -e 's/.*\.//' | sed -e 's/.*\///' | sort -u

如果你想要总数(有多少次扩展被看到):

find . -type f | sed -e 's/.*\.//' | sed -e 's/.*\///' | sort | uniq -c | sort -rn

非递归(单个文件夹):

for f in *.*; do printf "%s\n" "${f##*.}"; done | sort -u

我是根据这个论坛帖子写的，功劳应该在那里。

2009-12-03 19:38:02

Powershell：

dir -recurse | select-object extension -unique

感谢http://kevin-berridge.blogspot.com/2007/11/windows-powershell.html

2010-04-23 14:18:42

在Python中，为非常大的目录使用生成器，包括空白扩展名，并获取每个扩展名出现的次数:

import json
import collections
import itertools
import os

root = '/home/andres'
files = itertools.chain.from_iterable((
    files for _,_,files in os.walk(root)
    ))
counter = collections.Counter(
    (os.path.splitext(file_)[1] for file_ in files)
)
print json.dumps(counter, indent=2)

2012-08-24 19:17:28

如何在文件夹层次结构中找到所有不同的文件扩展名?

推荐文章

最新文章

标签