upstream/mercurial-mirror Files · hgext/convert/convcmd.py

compression: introduce a `storage.revlog.zstd.level` configuration...

compression: introduce a `storage.revlog.zstd.level` configuration This option control the zstd compression level used when compressing revlog chunk. The usage of zstd for revlog compression has not graduated from experimental yet, but we intend to fix that soon. The option name for the compression level is more straight forward to pick, so this changesets comes first. Having a dedicated option for each compression engine is useful because they don't support the same range of values. I ran the same measurement as for the zlib compression level (in the parent changesets). The variation in repository size is stay mostly in the same (small) range. The "read/write" performance see smallish variation, but are overall much better than zlib. Write performance show the same tend of having better write performance for when reaching high-end compression. Again, we don't intend to change the default zstd compression level (currently: 3) in this series. However this is worth investigating in the future. The Performance comparison of zlib vs zstd is quite impressive. The repository size stay in the same range, but the performance are much better in all situations. Comparison summary ================== We are looking at: - performance range for zlib - performance range for zstd - comparison of default zstd (level-3) to default zlib (level 6) - comparison of the slowest zstd time to the fastest zlib time Read performance: ----------------- | zlib | zstd | cmp | f2s mercurial | 0.170159 - 0.189219 | 0.144127 - 0.149624 | 80% | 88% pypy | 2.679217 - 2.768691 | 1.532317 - 1.705044 | 60% | 63% netbeans | 122.477027 - 141.620281 | 72.996346 - 89.731560 | 58% | 73% mozilla | 147.867662 - 170.572118 | 91.700995 - 105.853099 | 56% | 71% Write performance: ------------------ | zlib | zstd | cmp | f2s mercurial | 53.250304 - 56.2936129 | 40.877025 - 45.677286 | 75% | 86% pypy | 460.721984 - 476.589918 | 270.545409 - 301.002219 | 63% | 65% netbeans | 520.560316 - 715.930400 | 370.356311 - 428.329652 | 55% | 82% mozilla | 739.803002 - 987.056093 | 505.152906 - 591.930683 | 57% | 80% Raw data -------- repo alg lvl .hg/store size 00manifest.d read write mercurial zlib 1 49,402,813 5,963,475 0.170159 53.250304 mercurial zlib 6 47,197,397 5,875,730 0.182820 56.264320 mercurial zlib 9 47,121,596 5,849,781 0.189219 56.293612 mercurial zstd 1 49,737,084 5,966,355 0.144127 40.877025 mercurial zstd 3 48,961,867 5,895,208 0.146376 42.268142 mercurial zstd 5 48,200,592 5,938,676 0.149624 43.162875 mercurial zstd 10 47,833,520 5,913,353 0.145185 44.012489 mercurial zstd 15 47,314,604 5,728,679 0.147686 45.677286 mercurial zstd 20 47,330,502 5,830,539 0.145789 45.025407 mercurial zstd 22 47,330,076 5,830,539 0.143996 44.690460 pypy zlib 1 370,830,572 28,462,425 2.679217 460.721984 pypy zlib 6 340,112,317 27,648,747 2.768691 467.537158 pypy zlib 9 338,360,736 27,639,003 2.763495 476.589918 pypy zstd 1 362,377,479 27,916,214 1.532317 270.545409 pypy zstd 3 354,137,693 27,905,988 1.686718 294.951509 pypy zstd 5 342,640,043 27,655,774 1.705044 301.002219 pypy zstd 10 334,224,327 27,164,493 1.567287 285.186239 pypy zstd 15 329,000,363 26,645,965 1.637729 299.561332 pypy zstd 20 324,534,039 26,199,547 1.526813 302.149827 pypy zstd 22 324,530,595 26,198,932 1.525718 307.821218 netbeans zlib 1 1,281,847,810 165,495,457 122.477027 520.560316 netbeans zlib 6 1,205,284,353 159,161,207 139.876147 715.930400 netbeans zlib 9 1,197,135,671 155,034,586 141.620281 678.297064 netbeans zstd 1 1,259,581,737 160,840,613 72.996346 370.356311 netbeans zstd 3 1,232,978,122 157,691,551 81.622317 396.733087 netbeans zstd 5 1,208,034,075 160,246,880 83.080549 364.342626 netbeans zstd 10 1,188,624,176 156,083,417 79.323935 403.594602 netbeans zstd 15 1,176,973,589 153,859,477 89.731560 428.329652 netbeans zstd 20 1,162,958,258 151,147,535 82.842667 392.335349 netbeans zstd 22 1,162,707,029 151,150,220 82.565695 402.840655 mozilla zlib 1 2,775,497,186 298,527,987 147.867662 751.263721 mozilla zlib 6 2,596,856,420 286,597,671 170.572118 987.056093 mozilla zlib 9 2,587,542,494 287,018,264 163.622338 739.803002 mozilla zstd 1 2,723,159,348 286,617,532 91.700995 570.042751 mozilla zstd 3 2,665,055,001 286,152,013 95.240155 561.412805 mozilla zstd 5 2,607,819,817 288,060,030 101.978048 505.152906 mozilla zstd 10 2,558,761,085 283,967,648 104.113481 497.771202 mozilla zstd 15 2,526,216,060 275,581,300 105.853099 591.930683 mozilla zstd 20 2,485,114,806 266,478,859 95.268795 576.515389 mozilla zstd 22 2,484,869,080 266,456,505 94.429282 572.785537

Gregory Szorc - - Load All Authors

File last commit:

r41463:26761665 default


                r42211:bb271ec2

default

Download file

             convcmd.py
        
                    616 lines
            
             | 21.4 KiB
            
                | text/x-python
            
             |
                PythonLexer
            
             / hgext / convert / convcmd.py
          
                    History
                
                 |
                  Annotation
                 | Raw
                 |Copy content
                 |Copy permalink

      # convcmd - convert extension commands definition

      #

      # Copyright 2005-2007 Matt Mackall <mpm@selenic.com>

      #

      # This software may be used and distributed according to the terms of the

      # GNU General Public License version 2 or any later version.

      from __future__ import absolute_import

      import collections

      import os

      import shutil

      from mercurial.i18n import _

      from mercurial import (

          encoding,

          error,

          hg,

          pycompat,

          scmutil,

          util,

      )

      from mercurial.utils import dateutil

      from . import (

          bzr,

          common,

          cvs,

          darcs,

          filemap,

          git,

          gnuarch,

          hg as hgconvert,

          monotone,

          p4,

          subversion,

      )

      mapfile = common.mapfile

      MissingTool = common.MissingTool

      NoRepo = common.NoRepo

      SKIPREV = common.SKIPREV

      bzr_source = bzr.bzr_source

      convert_cvs = cvs.convert_cvs

      convert_git = git.convert_git

      darcs_source = darcs.darcs_source

      gnuarch_source = gnuarch.gnuarch_source

      mercurial_sink = hgconvert.mercurial_sink

      mercurial_source = hgconvert.mercurial_source

      monotone_source = monotone.monotone_source

      p4_source = p4.p4_source

      svn_sink = subversion.svn_sink

      svn_source = subversion.svn_source

      orig_encoding = 'ascii'

      def recode(s):

          if isinstance(s, pycompat.unicode):

              return s.encode(pycompat.sysstr(orig_encoding), 'replace')

          else:

              return s.decode('utf-8').encode(

                  pycompat.sysstr(orig_encoding), 'replace')

      def mapbranch(branch, branchmap):

          '''

          >>> bmap = {b'default': b'branch1'}

          >>> for i in [b'', None]:

          ...     mapbranch(i, bmap)

          'branch1'

          'branch1'

          >>> bmap = {b'None': b'branch2'}

          >>> for i in [b'', None]:

          ...     mapbranch(i, bmap)

          'branch2'

          'branch2'

          >>> bmap = {b'None': b'branch3', b'default': b'branch4'}

          >>> for i in [b'None', b'', None, b'default', b'branch5']:

          ...     mapbranch(i, bmap)

          'branch3'

          'branch4'

          'branch4'

          'branch4'

          'branch5'

          '''

          # If branch is None or empty, this commit is coming from the source

          # repository's default branch and destined for the default branch in the

          # destination repository. For such commits, using a literal "default"

          # in branchmap below allows the user to map "default" to an alternate

          # default branch in the destination repository.

          branch = branchmap.get(branch or 'default', branch)

          # At some point we used "None" literal to denote the default branch,

          # attempt to use that for backward compatibility.

          if (not branch):

              branch = branchmap.get('None', branch)

          return branch

      source_converters = [

          ('cvs', convert_cvs, 'branchsort'),

          ('git', convert_git, 'branchsort'),

          ('svn', svn_source, 'branchsort'),

          ('hg', mercurial_source, 'sourcesort'),

          ('darcs', darcs_source, 'branchsort'),

          ('mtn', monotone_source, 'branchsort'),

          ('gnuarch', gnuarch_source, 'branchsort'),

          ('bzr', bzr_source, 'branchsort'),

          ('p4', p4_source, 'branchsort'),

          ]

      sink_converters = [

          ('hg', mercurial_sink),

          ('svn', svn_sink),

          ]

      def convertsource(ui, path, type, revs):

          exceptions = []

          if type and type not in [s[0] for s in source_converters]:

              raise error.Abort(_('%s: invalid source repository type') % type)

          for name, source, sortmode in source_converters:

              try:

                  if not type or name == type:

                      return source(ui, name, path, revs), sortmode

              except (NoRepo, MissingTool) as inst:

                  exceptions.append(inst)

          if not ui.quiet:

              for inst in exceptions:

                  ui.write("%s\n" % pycompat.bytestr(inst.args[0]))

          raise error.Abort(_('%s: missing or unsupported repository') % path)

      def convertsink(ui, path, type):

          if type and type not in [s[0] for s in sink_converters]:

              raise error.Abort(_('%s: invalid destination repository type') % type)

          for name, sink in sink_converters:

              try:

                  if not type or name == type:

                      return sink(ui, name, path)

              except NoRepo as inst:

                  ui.note(_("convert: %s\n") % inst)

              except MissingTool as inst:

                  raise error.Abort('%s\n' % inst)

          raise error.Abort(_('%s: unknown repository type') % path)

      class progresssource(object):

          def __init__(self, ui, source, filecount):

              self.ui = ui

              self.source = source

              self.progress = ui.makeprogress(_('getting files'), unit=_('files'),

                                              total=filecount)

          def getfile(self, file, rev):

              self.progress.increment(item=file)

              return self.source.getfile(file, rev)

          def targetfilebelongstosource(self, targetfilename):

              return self.source.targetfilebelongstosource(targetfilename)

          def lookuprev(self, rev):

              return self.source.lookuprev(rev)

          def close(self):

              self.progress.complete()

      class converter(object):

          def __init__(self, ui, source, dest, revmapfile, opts):

              self.source = source

              self.dest = dest

              self.ui = ui

              self.opts = opts

              self.commitcache = {}

              self.authors = {}

              self.authorfile = None

              # Record converted revisions persistently: maps source revision

              # ID to target revision ID (both strings).  (This is how

              # incremental conversions work.)

              self.map = mapfile(ui, revmapfile)

              # Read first the dst author map if any

              authorfile = self.dest.authorfile()

              if authorfile and os.path.exists(authorfile):

                  self.readauthormap(authorfile)

              # Extend/Override with new author map if necessary

              if opts.get('authormap'):

                  self.readauthormap(opts.get('authormap'))

                  self.authorfile = self.dest.authorfile()

              self.splicemap = self.parsesplicemap(opts.get('splicemap'))

              self.branchmap = mapfile(ui, opts.get('branchmap'))

          def parsesplicemap(self, path):

              """ check and validate the splicemap format and

                  return a child/parents dictionary.

                  Format checking has two parts.

                  1. generic format which is same across all source types

                  2. specific format checking which may be different for

                     different source type.  This logic is implemented in

                     checkrevformat function in source files like

                     hg.py, subversion.py etc.

              """

              if not path:

                  return {}

              m = {}

              try:

                  fp = open(path, 'rb')

                  for i, line in enumerate(util.iterfile(fp)):

                      line = line.splitlines()[0].rstrip()

                      if not line:

                          # Ignore blank lines

                          continue

                      # split line

                      lex = common.shlexer(data=line, whitespace=',')

                      line = list(lex)

                      # check number of parents

                      if not (2 <= len(line) <= 3):

                          raise error.Abort(_('syntax error in %s(%d): child parent1'

                                             '[,parent2] expected') % (path, i + 1))

                      for part in line:

                          self.source.checkrevformat(part)

                      child, p1, p2 = line[0], line[1:2], line[2:]

                      if p1 == p2:

                          m[child] = p1

                      else:

                          m[child] = p1 + p2

               # if file does not exist or error reading, exit

              except IOError:

                  raise error.Abort(_('splicemap file not found or error reading %s:')

                                     % path)

              return m

          def walktree(self, heads):

              '''Return a mapping that identifies the uncommitted parents of every

              uncommitted changeset.'''

              visit = list(heads)

              known = set()

              parents = {}

              numcommits = self.source.numcommits()

              progress = self.ui.makeprogress(_('scanning'), unit=_('revisions'),

                                              total=numcommits)

              while visit:

                  n = visit.pop(0)

                  if n in known:

                      continue

                  if n in self.map:

                      m = self.map[n]

                      if m == SKIPREV or self.dest.hascommitfrommap(m):

                          continue

                  known.add(n)

                  progress.update(len(known))

                  commit = self.cachecommit(n)

                  parents[n] = []

                  for p in commit.parents:

                      parents[n].append(p)

                      visit.append(p)

              progress.complete()

              return parents

          def mergesplicemap(self, parents, splicemap):

              """A splicemap redefines child/parent relationships. Check the

              map contains valid revision identifiers and merge the new

              links in the source graph.

              """

              for c in sorted(splicemap):

                  if c not in parents:

                      if not self.dest.hascommitforsplicemap(self.map.get(c, c)):

                          # Could be in source but not converted during this run

                          self.ui.warn(_('splice map revision %s is not being '

                                         'converted, ignoring\n') % c)

                      continue

                  pc = []

                  for p in splicemap[c]:

                      # We do not have to wait for nodes already in dest.

                      if self.dest.hascommitforsplicemap(self.map.get(p, p)):

                          continue

                      # Parent is not in dest and not being converted, not good

                      if p not in parents:

                          raise error.Abort(_('unknown splice map parent: %s') % p)

                      pc.append(p)

                  parents[c] = pc

          def toposort(self, parents, sortmode):

              '''Return an ordering such that every uncommitted changeset is

              preceded by all its uncommitted ancestors.'''

              def mapchildren(parents):

                  """Return a (children, roots) tuple where 'children' maps parent

                  revision identifiers to children ones, and 'roots' is the list of

                  revisions without parents. 'parents' must be a mapping of revision

                  identifier to its parents ones.

                  """

                  visit = collections.deque(sorted(parents))

                  seen = set()

                  children = {}

                  roots = []

                  while visit:

                      n = visit.popleft()

                      if n in seen:

                          continue

                      seen.add(n)

                      # Ensure that nodes without parents are present in the

                      # 'children' mapping.

                      children.setdefault(n, [])

                      hasparent = False

                      for p in parents[n]:

                          if p not in self.map:

                              visit.append(p)

                              hasparent = True

                          children.setdefault(p, []).append(n)

                      if not hasparent:

                          roots.append(n)

                  return children, roots

              # Sort functions are supposed to take a list of revisions which

              # can be converted immediately and pick one

              def makebranchsorter():

                  """If the previously converted revision has a child in the

                  eligible revisions list, pick it. Return the list head

                  otherwise. Branch sort attempts to minimize branch

                  switching, which is harmful for Mercurial backend

                  compression.

                  """

                  prev = [None]

                  def picknext(nodes):

                      next = nodes[0]

                      for n in nodes:

                          if prev[0] in parents[n]:

                              next = n

                              break

                      prev[0] = next

                      return next

                  return picknext

              def makesourcesorter():

                  """Source specific sort."""

                  keyfn = lambda n: self.commitcache[n].sortkey

                  def picknext(nodes):

                      return sorted(nodes, key=keyfn)[0]

                  return picknext

              def makeclosesorter():

                  """Close order sort."""

                  keyfn = lambda n: ('close' not in self.commitcache[n].extra,

                                     self.commitcache[n].sortkey)

                  def picknext(nodes):

                      return sorted(nodes, key=keyfn)[0]

                  return picknext

              def makedatesorter():

                  """Sort revisions by date."""

                  dates = {}

                  def getdate(n):

                      if n not in dates:

                          dates[n] = dateutil.parsedate(self.commitcache[n].date)

                      return dates[n]

                  def picknext(nodes):

                      return min([(getdate(n), n) for n in nodes])[1]

                  return picknext

              if sortmode == 'branchsort':

                  picknext = makebranchsorter()

              elif sortmode == 'datesort':

                  picknext = makedatesorter()

              elif sortmode == 'sourcesort':

                  picknext = makesourcesorter()

              elif sortmode == 'closesort':

                  picknext = makeclosesorter()

              else:

                  raise error.Abort(_('unknown sort mode: %s') % sortmode)

              children, actives = mapchildren(parents)

              s = []

              pendings = {}

              while actives:

                  n = picknext(actives)

                  actives.remove(n)

                  s.append(n)

                  # Update dependents list

                  for c in children.get(n, []):

                      if c not in pendings:

                          pendings[c] = [p for p in parents[c] if p not in self.map]

                      try:

                          pendings[c].remove(n)

                      except ValueError:

                          raise error.Abort(_('cycle detected between %s and %s')

                                             % (recode(c), recode(n)))

                      if not pendings[c]:

                          # Parents are converted, node is eligible

                          actives.insert(0, c)

                          pendings[c] = None

              if len(s) != len(parents):

                  raise error.Abort(_("not all revisions were sorted"))

              return s

          def writeauthormap(self):

              authorfile = self.authorfile

              if authorfile:

                  self.ui.status(_('writing author map file %s\n') % authorfile)

                  ofile = open(authorfile, 'wb+')

                  for author in self.authors:

                      ofile.write(util.tonativeeol("%s=%s\n"

                                                   % (author, self.authors[author])))

                  ofile.close()

          def readauthormap(self, authorfile):

              afile = open(authorfile, 'rb')

              for line in afile:

                  line = line.strip()

                  if not line or line.startswith('#'):

                      continue

                  try:

                      srcauthor, dstauthor = line.split('=', 1)

                  except ValueError:

                      msg = _('ignoring bad line in author map file %s: %s\n')

                      self.ui.warn(msg % (authorfile, line.rstrip()))

                      continue

                  srcauthor = srcauthor.strip()

                  dstauthor = dstauthor.strip()

                  if self.authors.get(srcauthor) in (None, dstauthor):

                      msg = _('mapping author %s to %s\n')

                      self.ui.debug(msg % (srcauthor, dstauthor))

                      self.authors[srcauthor] = dstauthor

                      continue

                  m = _('overriding mapping for author %s, was %s, will be %s\n')

                  self.ui.status(m % (srcauthor, self.authors[srcauthor], dstauthor))

              afile.close()

          def cachecommit(self, rev):

              commit = self.source.getcommit(rev)

              commit.author = self.authors.get(commit.author, commit.author)

              commit.branch = mapbranch(commit.branch, self.branchmap)

              self.commitcache[rev] = commit

              return commit

          def copy(self, rev):

              commit = self.commitcache[rev]

              full = self.opts.get('full')

              changes = self.source.getchanges(rev, full)

              if isinstance(changes, bytes):

                  if changes == SKIPREV:

                      dest = SKIPREV

                  else:

                      dest = self.map[changes]

                  self.map[rev] = dest

                  return

              files, copies, cleanp2 = changes

              pbranches = []

              if commit.parents:

                  for prev in commit.parents:

                      if prev not in self.commitcache:

                          self.cachecommit(prev)

                      pbranches.append((self.map[prev],

                                        self.commitcache[prev].branch))

              self.dest.setbranch(commit.branch, pbranches)

              try:

                  parents = self.splicemap[rev]

                  self.ui.status(_('spliced in %s as parents of %s\n') %

                                 (_(' and ').join(parents), rev))

                  parents = [self.map.get(p, p) for p in parents]

              except KeyError:

                  parents = [b[0] for b in pbranches]

                  parents.extend(self.map[x]

                                 for x in commit.optparents

                                 if x in self.map)

              if len(pbranches) != 2:

                  cleanp2 = set()

              if len(parents) < 3:

                  source = progresssource(self.ui, self.source, len(files))

              else:

                  # For an octopus merge, we end up traversing the list of

                  # changed files N-1 times. This tweak to the number of

                  # files makes it so the progress bar doesn't overflow

                  # itself.

                  source = progresssource(self.ui, self.source,

                                          len(files) * (len(parents) - 1))

              newnode = self.dest.putcommit(files, copies, parents, commit,

                                            source, self.map, full, cleanp2)

              source.close()

              self.source.converted(rev, newnode)

              self.map[rev] = newnode

          def convert(self, sortmode):

              try:

                  self.source.before()

                  self.dest.before()

                  self.source.setrevmap(self.map)

                  self.ui.status(_("scanning source...\n"))

                  heads = self.source.getheads()

                  parents = self.walktree(heads)

                  self.mergesplicemap(parents, self.splicemap)

                  self.ui.status(_("sorting...\n"))

                  t = self.toposort(parents, sortmode)

                  num = len(t)

                  c = None

                  self.ui.status(_("converting...\n"))

                  progress = self.ui.makeprogress(_('converting'),

                                                  unit=_('revisions'), total=len(t))

                  for i, c in enumerate(t):

                      num -= 1

                      desc = self.commitcache[c].desc

                      if "\n" in desc:

                          desc = desc.splitlines()[0]

                      # convert log message to local encoding without using

                      # tolocal() because the encoding.encoding convert()

                      # uses is 'utf-8'

                      self.ui.status("%d %s\n" % (num, recode(desc)))

                      self.ui.note(_("source: %s\n") % recode(c))

                      progress.update(i)

                      self.copy(c)

                  progress.complete()

                  if not self.ui.configbool('convert', 'skiptags'):

                      tags = self.source.gettags()

                      ctags = {}

                      for k in tags:

                          v = tags[k]

                          if self.map.get(v, SKIPREV) != SKIPREV:

                              ctags[k] = self.map[v]

                      if c and ctags:

                          nrev, tagsparent = self.dest.puttags(ctags)

                          if nrev and tagsparent:

                              # write another hash correspondence to override the

                              # previous one so we don't end up with extra tag heads

                              tagsparents = [e for e in self.map.iteritems()

                                             if e[1] == tagsparent]

                              if tagsparents:

                                  self.map[tagsparents[0][0]] = nrev

                  bookmarks = self.source.getbookmarks()

                  cbookmarks = {}

                  for k in bookmarks:

                      v = bookmarks[k]

                      if self.map.get(v, SKIPREV) != SKIPREV:

                          cbookmarks[k] = self.map[v]

                  if c and cbookmarks:

                      self.dest.putbookmarks(cbookmarks)

                  self.writeauthormap()

              finally:

                  self.cleanup()

          def cleanup(self):

              try:

                  self.dest.after()

              finally:

                  self.source.after()

              self.map.close()

      def convert(ui, src, dest=None, revmapfile=None, **opts):

          opts = pycompat.byteskwargs(opts)

          global orig_encoding

          orig_encoding = encoding.encoding

          encoding.encoding = 'UTF-8'

          # support --authors as an alias for --authormap

          if not opts.get('authormap'):

              opts['authormap'] = opts.get('authors')

          if not dest:

              dest = hg.defaultdest(src) + "-hg"

              ui.status(_("assuming destination %s\n") % dest)

          destc = convertsink(ui, dest, opts.get('dest_type'))

          destc = scmutil.wrapconvertsink(destc)

          try:

              srcc, defaultsort = convertsource(ui, src, opts.get('source_type'),

                                                opts.get('rev'))

          except Exception:

              for path in destc.created:

                  shutil.rmtree(path, True)

              raise

          sortmodes = ('branchsort', 'datesort', 'sourcesort', 'closesort')

          sortmode = [m for m in sortmodes if opts.get(m)]

          if len(sortmode) > 1:

              raise error.Abort(_('more than one sort mode specified'))

          if sortmode:

              sortmode = sortmode[0]

          else:

              sortmode = defaultsort

          if sortmode == 'sourcesort' and not srcc.hasnativeorder():

              raise error.Abort(_('--sourcesort is not supported by this data source')

                               )

          if sortmode == 'closesort' and not srcc.hasnativeclose():

              raise error.Abort(_('--closesort is not supported by this data source'))

          fmap = opts.get('filemap')

          if fmap:

              srcc = filemap.filemap_source(ui, srcc, fmap)

              destc.setfilemapmode(True)

          if not revmapfile:

              revmapfile = destc.revmapfile()

          c = converter(ui, srcc, destc, revmapfile, opts)

          c.convert(sortmode)

	Site-wide shortcuts
/	Use quick search box
g h	Goto home page
g g	Goto my private gists page
g G	Goto my public gists page
g 0-9	Goto bookmarked items from 0-9
n r	New repository page
n g	New gist page

	Repositories
g s	Goto summary page
g c	Goto changelog page
g f	Goto files page
g F	Goto files page with file search activated
g p	Goto pull requests page
g o	Goto repository settings
g O	Goto repository access permissions settings
t s	Toggle sidebar on some pages

				# convcmd - convert extension commands definition
				#
				# Copyright 2005-2007 Matt Mackall <mpm@selenic.com>
				#
				# This software may be used and distributed according to the terms of the
				# GNU General Public License version 2 or any later version.
				from __future__ import absolute_import

				import collections
				import os
				import shutil

				from mercurial.i18n import _
				from mercurial import (
				encoding,
				error,
				hg,
				pycompat,
				scmutil,
				util,
				)
				from mercurial.utils import dateutil

				from . import (
				bzr,
				common,
				cvs,
				darcs,
				filemap,
				git,
				gnuarch,
				hg as hgconvert,
				monotone,
				p4,
				subversion,
				)

				mapfile = common.mapfile
				MissingTool = common.MissingTool
				NoRepo = common.NoRepo
				SKIPREV = common.SKIPREV

				bzr_source = bzr.bzr_source
				convert_cvs = cvs.convert_cvs
				convert_git = git.convert_git
				darcs_source = darcs.darcs_source
				gnuarch_source = gnuarch.gnuarch_source
				mercurial_sink = hgconvert.mercurial_sink
				mercurial_source = hgconvert.mercurial_source
				monotone_source = monotone.monotone_source
				p4_source = p4.p4_source
				svn_sink = subversion.svn_sink
				svn_source = subversion.svn_source

				orig_encoding = 'ascii'

				def recode(s):
				if isinstance(s, pycompat.unicode):
				return s.encode(pycompat.sysstr(orig_encoding), 'replace')
				else:
				return s.decode('utf-8').encode(
				pycompat.sysstr(orig_encoding), 'replace')

				def mapbranch(branch, branchmap):
				'''
				>>> bmap = {b'default': b'branch1'}
				>>> for i in [b'', None]:
				... mapbranch(i, bmap)
				'branch1'
				'branch1'
				>>> bmap = {b'None': b'branch2'}
				>>> for i in [b'', None]:
				... mapbranch(i, bmap)
				'branch2'
				'branch2'
				>>> bmap = {b'None': b'branch3', b'default': b'branch4'}
				>>> for i in [b'None', b'', None, b'default', b'branch5']:
				... mapbranch(i, bmap)
				'branch3'
				'branch4'
				'branch4'
				'branch4'
				'branch5'
				'''
				# If branch is None or empty, this commit is coming from the source
				# repository's default branch and destined for the default branch in the
				# destination repository. For such commits, using a literal "default"
				# in branchmap below allows the user to map "default" to an alternate
				# default branch in the destination repository.
				branch = branchmap.get(branch or 'default', branch)
				# At some point we used "None" literal to denote the default branch,
				# attempt to use that for backward compatibility.
				if (not branch):
				branch = branchmap.get('None', branch)
				return branch

				source_converters = [
				('cvs', convert_cvs, 'branchsort'),
				('git', convert_git, 'branchsort'),
				('svn', svn_source, 'branchsort'),
				('hg', mercurial_source, 'sourcesort'),
				('darcs', darcs_source, 'branchsort'),
				('mtn', monotone_source, 'branchsort'),
				('gnuarch', gnuarch_source, 'branchsort'),
				('bzr', bzr_source, 'branchsort'),
				('p4', p4_source, 'branchsort'),
				]

				sink_converters = [
				('hg', mercurial_sink),
				('svn', svn_sink),
				]

				def convertsource(ui, path, type, revs):
				exceptions = []
				if type and type not in [s[0] for s in source_converters]:
				raise error.Abort(_('%s: invalid source repository type') % type)
				for name, source, sortmode in source_converters:
				try:
				if not type or name == type:
				return source(ui, name, path, revs), sortmode
				except (NoRepo, MissingTool) as inst:
				exceptions.append(inst)
				if not ui.quiet:
				for inst in exceptions:
				ui.write("%s\n" % pycompat.bytestr(inst.args[0]))
				raise error.Abort(_('%s: missing or unsupported repository') % path)

				def convertsink(ui, path, type):
				if type and type not in [s[0] for s in sink_converters]:
				raise error.Abort(_('%s: invalid destination repository type') % type)
				for name, sink in sink_converters:
				try:
				if not type or name == type:
				return sink(ui, name, path)
				except NoRepo as inst:
				ui.note(_("convert: %s\n") % inst)
				except MissingTool as inst:
				raise error.Abort('%s\n' % inst)
				raise error.Abort(_('%s: unknown repository type') % path)

				class progresssource(object):
				def __init__(self, ui, source, filecount):
				self.ui = ui
				self.source = source
				self.progress = ui.makeprogress(_('getting files'), unit=_('files'),
				total=filecount)

				def getfile(self, file, rev):
				self.progress.increment(item=file)
				return self.source.getfile(file, rev)

				def targetfilebelongstosource(self, targetfilename):
				return self.source.targetfilebelongstosource(targetfilename)

				def lookuprev(self, rev):
				return self.source.lookuprev(rev)

				def close(self):
				self.progress.complete()

				class converter(object):
				def __init__(self, ui, source, dest, revmapfile, opts):

				self.source = source
				self.dest = dest
				self.ui = ui
				self.opts = opts
				self.commitcache = {}
				self.authors = {}
				self.authorfile = None

				# Record converted revisions persistently: maps source revision
				# ID to target revision ID (both strings). (This is how
				# incremental conversions work.)
				self.map = mapfile(ui, revmapfile)

				# Read first the dst author map if any
				authorfile = self.dest.authorfile()
				if authorfile and os.path.exists(authorfile):
				self.readauthormap(authorfile)
				# Extend/Override with new author map if necessary
				if opts.get('authormap'):
				self.readauthormap(opts.get('authormap'))
				self.authorfile = self.dest.authorfile()

				self.splicemap = self.parsesplicemap(opts.get('splicemap'))
				self.branchmap = mapfile(ui, opts.get('branchmap'))

				def parsesplicemap(self, path):
				""" check and validate the splicemap format and
				return a child/parents dictionary.
				Format checking has two parts.
				1. generic format which is same across all source types
				2. specific format checking which may be different for
				different source type. This logic is implemented in
				checkrevformat function in source files like
				hg.py, subversion.py etc.
				"""

				if not path:
				return {}
				m = {}
				try:
				fp = open(path, 'rb')
				for i, line in enumerate(util.iterfile(fp)):
				line = line.splitlines()[0].rstrip()
				if not line:
				# Ignore blank lines
				continue
				# split line
				lex = common.shlexer(data=line, whitespace=',')
				line = list(lex)
				# check number of parents
				if not (2 <= len(line) <= 3):
				raise error.Abort(_('syntax error in %s(%d): child parent1'
				'[,parent2] expected') % (path, i + 1))
				for part in line:
				self.source.checkrevformat(part)
				child, p1, p2 = line[0], line[1:2], line[2:]
				if p1 == p2:
				m[child] = p1
				else:
				m[child] = p1 + p2
				# if file does not exist or error reading, exit
				except IOError:
				raise error.Abort(_('splicemap file not found or error reading %s:')
				% path)
				return m


				def walktree(self, heads):
				'''Return a mapping that identifies the uncommitted parents of every
				uncommitted changeset.'''
				visit = list(heads)
				known = set()
				parents = {}
				numcommits = self.source.numcommits()
				progress = self.ui.makeprogress(_('scanning'), unit=_('revisions'),
				total=numcommits)
				while visit:
				n = visit.pop(0)
				if n in known:
				continue
				if n in self.map:
				m = self.map[n]
				if m == SKIPREV or self.dest.hascommitfrommap(m):
				continue
				known.add(n)
				progress.update(len(known))
				commit = self.cachecommit(n)
				parents[n] = []
				for p in commit.parents:
				parents[n].append(p)
				visit.append(p)
				progress.complete()

				return parents

				def mergesplicemap(self, parents, splicemap):
				"""A splicemap redefines child/parent relationships. Check the
				map contains valid revision identifiers and merge the new
				links in the source graph.
				"""
				for c in sorted(splicemap):
				if c not in parents:
				if not self.dest.hascommitforsplicemap(self.map.get(c, c)):
				# Could be in source but not converted during this run
				self.ui.warn(_('splice map revision %s is not being '
				'converted, ignoring\n') % c)
				continue
				pc = []
				for p in splicemap[c]:
				# We do not have to wait for nodes already in dest.
				if self.dest.hascommitforsplicemap(self.map.get(p, p)):
				continue
				# Parent is not in dest and not being converted, not good
				if p not in parents:
				raise error.Abort(_('unknown splice map parent: %s') % p)
				pc.append(p)
				parents[c] = pc

				def toposort(self, parents, sortmode):
				'''Return an ordering such that every uncommitted changeset is
				preceded by all its uncommitted ancestors.'''

				def mapchildren(parents):
				"""Return a (children, roots) tuple where 'children' maps parent
				revision identifiers to children ones, and 'roots' is the list of
				revisions without parents. 'parents' must be a mapping of revision
				identifier to its parents ones.
				"""
				visit = collections.deque(sorted(parents))
				seen = set()
				children = {}
				roots = []

				while visit:
				n = visit.popleft()
				if n in seen:
				continue
				seen.add(n)
				# Ensure that nodes without parents are present in the
				# 'children' mapping.
				children.setdefault(n, [])
				hasparent = False
				for p in parents[n]:
				if p not in self.map:
				visit.append(p)
				hasparent = True
				children.setdefault(p, []).append(n)
				if not hasparent:
				roots.append(n)

				return children, roots

				# Sort functions are supposed to take a list of revisions which
				# can be converted immediately and pick one

				def makebranchsorter():
				"""If the previously converted revision has a child in the
				eligible revisions list, pick it. Return the list head
				otherwise. Branch sort attempts to minimize branch
				switching, which is harmful for Mercurial backend
				compression.
				"""
				prev = [None]
				def picknext(nodes):
				next = nodes[0]
				for n in nodes:
				if prev[0] in parents[n]:
				next = n
				break
				prev[0] = next
				return next
				return picknext

				def makesourcesorter():
				"""Source specific sort."""
				keyfn = lambda n: self.commitcache[n].sortkey
				def picknext(nodes):
				return sorted(nodes, key=keyfn)[0]
				return picknext

				def makeclosesorter():
				"""Close order sort."""
				keyfn = lambda n: ('close' not in self.commitcache[n].extra,
				self.commitcache[n].sortkey)
				def picknext(nodes):
				return sorted(nodes, key=keyfn)[0]
				return picknext

				def makedatesorter():
				"""Sort revisions by date."""
				dates = {}
				def getdate(n):
				if n not in dates:
				dates[n] = dateutil.parsedate(self.commitcache[n].date)
				return dates[n]

				def picknext(nodes):
				return min([(getdate(n), n) for n in nodes])[1]

				return picknext

				if sortmode == 'branchsort':
				picknext = makebranchsorter()
				elif sortmode == 'datesort':
				picknext = makedatesorter()
				elif sortmode == 'sourcesort':
				picknext = makesourcesorter()
				elif sortmode == 'closesort':
				picknext = makeclosesorter()
				else:
				raise error.Abort(_('unknown sort mode: %s') % sortmode)

				children, actives = mapchildren(parents)

				s = []
				pendings = {}
				while actives:
				n = picknext(actives)
				actives.remove(n)
				s.append(n)

				# Update dependents list
				for c in children.get(n, []):
				if c not in pendings:
				pendings[c] = [p for p in parents[c] if p not in self.map]
				try:
				pendings[c].remove(n)
				except ValueError:
				raise error.Abort(_('cycle detected between %s and %s')
				% (recode(c), recode(n)))
				if not pendings[c]:
				# Parents are converted, node is eligible
				actives.insert(0, c)
				pendings[c] = None

				if len(s) != len(parents):
				raise error.Abort(_("not all revisions were sorted"))

				return s

				def writeauthormap(self):
				authorfile = self.authorfile
				if authorfile:
				self.ui.status(_('writing author map file %s\n') % authorfile)
				ofile = open(authorfile, 'wb+')
				for author in self.authors:
				ofile.write(util.tonativeeol("%s=%s\n"
				% (author, self.authors[author])))
				ofile.close()

				def readauthormap(self, authorfile):
				afile = open(authorfile, 'rb')
				for line in afile:

				line = line.strip()
				if not line or line.startswith('#'):
				continue

				try:
				srcauthor, dstauthor = line.split('=', 1)
				except ValueError:
				msg = _('ignoring bad line in author map file %s: %s\n')
				self.ui.warn(msg % (authorfile, line.rstrip()))
				continue

				srcauthor = srcauthor.strip()
				dstauthor = dstauthor.strip()
				if self.authors.get(srcauthor) in (None, dstauthor):
				msg = _('mapping author %s to %s\n')
				self.ui.debug(msg % (srcauthor, dstauthor))
				self.authors[srcauthor] = dstauthor
				continue

				m = _('overriding mapping for author %s, was %s, will be %s\n')
				self.ui.status(m % (srcauthor, self.authors[srcauthor], dstauthor))

				afile.close()

				def cachecommit(self, rev):
				commit = self.source.getcommit(rev)
				commit.author = self.authors.get(commit.author, commit.author)
				commit.branch = mapbranch(commit.branch, self.branchmap)
				self.commitcache[rev] = commit
				return commit

				def copy(self, rev):
				commit = self.commitcache[rev]
				full = self.opts.get('full')
				changes = self.source.getchanges(rev, full)
				if isinstance(changes, bytes):
				if changes == SKIPREV:
				dest = SKIPREV
				else:
				dest = self.map[changes]
				self.map[rev] = dest
				return
				files, copies, cleanp2 = changes
				pbranches = []
				if commit.parents:
				for prev in commit.parents:
				if prev not in self.commitcache:
				self.cachecommit(prev)
				pbranches.append((self.map[prev],
				self.commitcache[prev].branch))
				self.dest.setbranch(commit.branch, pbranches)
				try:
				parents = self.splicemap[rev]
				self.ui.status(_('spliced in %s as parents of %s\n') %
				(_(' and ').join(parents), rev))
				parents = [self.map.get(p, p) for p in parents]
				except KeyError:
				parents = [b[0] for b in pbranches]
				parents.extend(self.map[x]
				for x in commit.optparents
				if x in self.map)
				if len(pbranches) != 2:
				cleanp2 = set()
				if len(parents) < 3:
				source = progresssource(self.ui, self.source, len(files))
				else:
				# For an octopus merge, we end up traversing the list of
				# changed files N-1 times. This tweak to the number of
				# files makes it so the progress bar doesn't overflow
				# itself.
				source = progresssource(self.ui, self.source,
				len(files) * (len(parents) - 1))
				newnode = self.dest.putcommit(files, copies, parents, commit,
				source, self.map, full, cleanp2)
				source.close()
				self.source.converted(rev, newnode)
				self.map[rev] = newnode

				def convert(self, sortmode):
				try:
				self.source.before()
				self.dest.before()
				self.source.setrevmap(self.map)
				self.ui.status(_("scanning source...\n"))
				heads = self.source.getheads()
				parents = self.walktree(heads)
				self.mergesplicemap(parents, self.splicemap)
				self.ui.status(_("sorting...\n"))
				t = self.toposort(parents, sortmode)
				num = len(t)
				c = None

				self.ui.status(_("converting...\n"))
				progress = self.ui.makeprogress(_('converting'),
				unit=_('revisions'), total=len(t))
				for i, c in enumerate(t):
				num -= 1
				desc = self.commitcache[c].desc
				if "\n" in desc:
				desc = desc.splitlines()[0]
				# convert log message to local encoding without using
				# tolocal() because the encoding.encoding convert()
				# uses is 'utf-8'
				self.ui.status("%d %s\n" % (num, recode(desc)))
				self.ui.note(_("source: %s\n") % recode(c))
				progress.update(i)
				self.copy(c)
				progress.complete()

				if not self.ui.configbool('convert', 'skiptags'):
				tags = self.source.gettags()
				ctags = {}
				for k in tags:
				v = tags[k]
				if self.map.get(v, SKIPREV) != SKIPREV:
				ctags[k] = self.map[v]

				if c and ctags:
				nrev, tagsparent = self.dest.puttags(ctags)
				if nrev and tagsparent:
				# write another hash correspondence to override the
				# previous one so we don't end up with extra tag heads
				tagsparents = [e for e in self.map.iteritems()
				if e[1] == tagsparent]
				if tagsparents:
				self.map[tagsparents[0][0]] = nrev

				bookmarks = self.source.getbookmarks()
				cbookmarks = {}
				for k in bookmarks:
				v = bookmarks[k]
				if self.map.get(v, SKIPREV) != SKIPREV:
				cbookmarks[k] = self.map[v]

				if c and cbookmarks:
				self.dest.putbookmarks(cbookmarks)

				self.writeauthormap()
				finally:
				self.cleanup()

				def cleanup(self):
				try:
				self.dest.after()
				finally:
				self.source.after()
				self.map.close()

				def convert(ui, src, dest=None, revmapfile=None, **opts):
				opts = pycompat.byteskwargs(opts)
				global orig_encoding
				orig_encoding = encoding.encoding
				encoding.encoding = 'UTF-8'

				# support --authors as an alias for --authormap
				if not opts.get('authormap'):
				opts['authormap'] = opts.get('authors')

				if not dest:
				dest = hg.defaultdest(src) + "-hg"
				ui.status(_("assuming destination %s\n") % dest)

				destc = convertsink(ui, dest, opts.get('dest_type'))
				destc = scmutil.wrapconvertsink(destc)

				try:
				srcc, defaultsort = convertsource(ui, src, opts.get('source_type'),
				opts.get('rev'))
				except Exception:
				for path in destc.created:
				shutil.rmtree(path, True)
				raise

				sortmodes = ('branchsort', 'datesort', 'sourcesort', 'closesort')
				sortmode = [m for m in sortmodes if opts.get(m)]
				if len(sortmode) > 1:
				raise error.Abort(_('more than one sort mode specified'))
				if sortmode:
				sortmode = sortmode[0]
				else:
				sortmode = defaultsort

				if sortmode == 'sourcesort' and not srcc.hasnativeorder():
				raise error.Abort(_('--sourcesort is not supported by this data source')
				)
				if sortmode == 'closesort' and not srcc.hasnativeclose():
				raise error.Abort(_('--closesort is not supported by this data source'))

				fmap = opts.get('filemap')
				if fmap:
				srcc = filemap.filemap_source(ui, srcc, fmap)
				destc.setfilemapmode(True)

				if not revmapfile:
				revmapfile = destc.revmapfile()

				c = converter(ui, srcc, destc, revmapfile, opts)
				c.convert(sortmode)